SYSTEM all green source itma.com queue 2,419 pages p99 latency 218ms dataflirt.com · scraper/itma-com
RUN · 14 active pipelines · itma.com live

ITMA exhibitor data,
at warehouse scale.

We extract exhibitor profiles, product sectors, machinery classifications, booth mappings, and speaker schedules from ITMA. Delivered as clean JSON, CSV, or Parquet to S3, BigQuery, or Snowflake on your cadence.

Exhibitors extracted
1,842 /run
Product sectors
341 /run
Booth mappings
2,105 /run
Active pipelines
14
Uptime
99.94%
Data Dictionary

Every field we extract from itma.com

Structured, schema-consistent data across all major object types — delivered clean, typed, and ready to query.

Complete list of extractable fields for Exhibitor Profiles objects from itma.com. All fields typed and schema-versioned.

exhibitor_idcompany_namecountrywebsitedescriptionbooth_numberhall_numberproduct_sectorscontact_emailphone
exhibitor_profiles
● 200 OK
"exhibitor_id": "EX-8921",
"company_name": "Rieter Machine Works Ltd.",
"country": "Switzerland",
"website": "https://www.rieter.com",
"booth_number": "H2-C201",
"hall_number": "Hall 2",
"product_sectors": "['Spinning', 'Automation']",
"phone": "+41 52 208 7171"
# exhibitor_idcompany_namecountrywebsitedescriptionbooth_number
1
2
3

Complete list of extractable fields for Product Categories objects from itma.com. All fields typed and schema-versioned.

sector_idsector_nameparent_categorydescriptionexhibitor_countmachinery_typeapplication_areaurl
product_categories
● 200 OK
"sector_id": "CH-102",
"sector_name": "Ring Spinning Machines",
"parent_category": "Spinning",
"exhibitor_count": 42,
"machinery_type": "Yarn Production",
"application_area": "Apparel Textiles",
"url": "https://itma.com/sectors/spinning/ring-spinning"
# sector_idsector_nameparent_categorydescriptionexhibitor_countmachinery_type
1
2
3

Complete list of extractable fields for Hall & Booth Mapping objects from itma.com. All fields typed and schema-versioned.

hall_idhall_namefloor_levelbooth_idexhibitor_idbooth_size_sqmzone_typecoordinates
hall_& booth mapping
● 200 OK
"hall_id": "H2",
"hall_name": "Spinning & Winding",
"floor_level": "Ground",
"booth_id": "H2-C201",
"exhibitor_id": "EX-8921",
"zone_type": "Machinery",
"coordinates": "45.12, 9.18"
# hall_idhall_namefloor_levelbooth_idexhibitor_idbooth_size_sqm
1
2
3

Complete list of extractable fields for Events & Schedules objects from itma.com. All fields typed and schema-versioned.

event_idtitledatestart_timeend_timelocationspeaker_idstrackdescription
events_& schedules
● 200 OK
"event_id": "EV-304",
"title": "Sustainable Dyeing Innovations",
"date": "2027-06-10",
"start_time": "14:00",
"end_time": "15:30",
"location": "Innovation Forum, Hall 3",
"track": "Sustainability",
"speaker_ids": "['SPK-112', 'SPK-405']"
# event_idtitledatestart_timeend_timelocation
1
2
3

Complete list of extractable fields for Speakers objects from itma.com. All fields typed and schema-versioned.

speaker_idfull_namecompanyjob_titlebiographysession_idslinkedin_urlimage_url
speakers
● 200 OK
"speaker_id": "SPK-112",
"full_name": "Dr. Elena Rossi",
"company": "EcoTextile Research",
"job_title": "Head of Material Science",
"session_ids": "['EV-304']",
"linkedin_url": "https://linkedin.com/in/elenarossi",
"image_url": "https://itma.com/images/speakers/spk112.jpg"
# speaker_idfull_namecompanyjob_titlebiographysession_ids
1
2
3

Capabilities

Everything you need from ITMA, nothing you do not

Our ITMA scraper navigates the complex exhibitor directory, extracting nested machinery classifications, booth locations, and company profiles with full session management built in.

Exhibitor Directory Extraction

Extract company names, countries, booth numbers, and detailed descriptions across the entire ITMA exhibitor list.

Product Sector Mapping

Capture the complete taxonomy of textile machinery categories, from spinning and weaving to finishing and garment making.

Hall & Booth Intelligence

Map exhibitors to specific halls and booth sizes, providing spatial context for trade show logistics.

Contact Data Parsing

Extract public-facing website URLs, email addresses, and phone numbers for direct B2B outreach.

Innovation & Press Tracking

Monitor exhibitor press releases and innovation awards published on the ITMA portal.

Speaker & Conference Schedules

Extract session titles, times, locations, and speaker biographies from the ITMA conference agenda.

Country Pavilion Data

Group exhibitors by national associations and country pavilions for regional market analysis.

Change Detection

Track new exhibitor additions, booth reassignments, and cancellation updates leading up to the event.

Scheduled Export Modes

Configure one-off bulk extracts or regular weekly syncs to monitor directory updates.

// engagement pipeline

From directory to warehouse record

Brief in. Clean data out.

Define Scope
d 0

Provide target product sectors, halls, or specific exhibitor lists. We design the extraction schema together.

Pipeline Build
d 2–4

We configure Scrapy / Playwright crawlers, proxy rotation, and session management for itma.com.

Validation & QA
d 4–6

Schema validation, null-rate checks, and sample data review before full pipeline launch.

Delivery
ongoing

JSON / CSV / Parquet pushed to your S3 bucket, BigQuery dataset, or Snowflake stage on agreed cadence.

Under the hood

How our ITMA pipeline handles the hard parts

Trade show directories use aggressive rate limiting and complex JavaScript rendering to protect exhibitor data. Here is how we maintain extraction stability.

pipeline-monitor · itma.com · live ● active
// fingerprinting
Identity rotation
TLS fingerprintrandomised
User-agentrotated
IP poolresidential
Challenges blocked0
// pagination
Page coverage
48,291 pages queued running
// observability
Pipeline health
99.9%
uptime
142ms
p99 lat
0.3%
null rate
2
alerts
Anti-bot layer
Rate limiting and IP block evasion

Trade show directories implement strict request thresholds to prevent bulk scraping. We distribute requests across a pool of residential proxies, inserting randomised delays to mimic human browsing behaviour and avoid IP bans.

JavaScript rendering
Dynamic exhibitor lists

ITMA uses JavaScript frameworks to load exhibitor details and booth maps dynamically. Our pipeline uses Playwright to execute client-side scripts, ensuring we capture data that standard HTTP requests miss.

Schema stability
Handling DOM changes

Event portals frequently update their layout as the exhibition date approaches. We use resilient CSS and XPath selectors with multiple fallback chains to ensure continuous data flow.

Change detection
Monitoring directory updates

We maintain state across pipeline runs. When an exhibitor changes their booth number or adds a new product category, we emit only the updated records, saving compute and storage costs.

Monitoring & alerting
Anomaly detection

We monitor extraction metrics continuously. If the exhibitor count drops unexpectedly or pagination fails, our system alerts engineers immediately to adjust selectors.

Applications

Who uses ITMA data and how

Teams across industries use itma.com data to build competitive products and smarter operations.

01
Competitor Analysis

Textile machinery manufacturers track competitor booth sizes, product launches, and category positioning.

02
Lead Generation & Procurement

Textile mills and garment factories extract exhibitor contact details to schedule targeted B2B meetings.

03
Market Trend Mapping

Analysts track the growth of specific product sectors, such as sustainable dyeing or automation, by measuring exhibitor participation.

04
Event Planning & Logistics

Logistics providers use hall mapping and booth size data to target exhibitors needing freight and installation services.

05
B2B Matchmaking

Industry associations cross-reference ITMA exhibitor lists with their own member databases to facilitate trade delegations.

06
Investment Due Diligence

Private equity firms monitor the presence and scale of specific machinery manufacturers at the industry's premier event.

Why DataFlirt

"ITMA represents the global baseline for textile technology, but its directory is locked behind dynamic web views rather than queryable APIs."

Extracting exhibitor data requires navigating complex pagination, JavaScript-rendered booth maps, and strict rate limits. DataFlirt handles the proxy rotation, session management, and parsing logic so your team receives clean, structured data ready for analysis.

Technical Spec

ITMA scraper — technical capabilities

Everything supported by our itma.com scraper — rendered SPA elements, auth walls, rate-limit evasion and beyond.

JavaScript rendering
Full Playwright sessions required for dynamic exhibitor lists and maps
Supported
CAPTCHA bypass
Automated 2Captcha + CapSolver integration
Supported
Residential proxy rotation
ISP-grade residential IPs rotated to avoid rate limits
Supported
Exhibitor pagination
Traversal of all directory pages to ensure complete extraction
Supported
Nested product sectors
Extraction of hierarchical machinery classifications
Supported
Change detection
Hash-based diff to identify new or updated exhibitor profiles
Supported
Webhook delivery
HTTP POST per record or batch delivery options
Supported
Attendee list extraction
Private attendee data is strictly gated and not publicly accessible
Partial
ITMAconnect messaging
Direct B2B messaging requires authenticated user accounts
Partial
Infrastructure

Infrastructure powering the ITMA pipeline

Open-source tooling on proven cloud infra — no vendor lock-in, full observability.

ScrapyPlaywrightPython 3.12RedisPostgreSQLApache AirflowAWS LambdaS3CloudWatch2CaptchaCapSolverResidential ProxiesDockerKubernetesGrafanaPrometheus
Scrapy + Playwright Stack

Scrapy handles crawl orchestration and retry logic. Playwright handles JavaScript rendering and interaction flows for complex directory structures.

Residential Proxy Infrastructure

We maintain pools of residential ISP proxies. Rotation happens per-request to bypass strict rate limits imposed by event portals.

Cloud-Native Orchestration

Pipelines run on AWS Lambda and ECS. Airflow handles scheduling and dependency management, with all state stored in PostgreSQL.

Output & Delivery

Your data, your destination

Data delivered to where your team already works — no new tooling required.

JSON
Newline-delimited or nested array format
CSV
Flat file with typed columns
Parquet
Columnar format for data warehouses
S3
Direct bucket delivery
Webhook
HTTP POST per record
XLS
Excel compatible format for business users
API
REST endpoints to query extracted records
BigQuery
Streamed directly into your dataset
// faq

Common questions.

About itma.com scraping, legality, and pipeline operations.

Ask us directly →
Is scraping the ITMA directory legal?

Scraping publicly available information is generally permissible. DataFlirt extracts only public exhibitor and event data. We do not bypass authentication walls or extract private attendee information.

How often can you refresh the exhibitor list?

We typically run weekly or daily updates leading up to the exhibition. During the event, we can configure higher frequency runs if the platform publishes live updates.

Do you capture the full product sector hierarchy?

Yes. We extract parent categories, sub-categories, and specific machinery types, maintaining the relationships to each exhibitor.

Can you extract data from the interactive hall maps?

Yes. We parse the underlying data structures of interactive maps to extract hall numbers, booth IDs, and spatial coordinates where available.

Do you provide historical ITMA data?

We extract data currently live on the site. If archives of past exhibitions remain publicly accessible, we can target those specific URLs.

What format is best for importing into my CRM?

We recommend CSV or XLS for direct imports into Salesforce or HubSpot. We map the fields to match your target schema prior to delivery.

$ dataflirt scope --new-project --source=itma.com ready

Tell us what
to extract.
We do the rest.

20-minute scoping call. Pilot dataset within the week. Production within two. Whether you need a full exhibitor directory export or continuous monitoring of product sectors, we scope, build, and operate the pipeline. Tell us what you need.

hello@dataflirt.com · Bengaluru · IST · typical reply < 4h
Related Scrapers

More in textile and fabric

Services

Data Extraction for Every Industry

View All Services →