We extract luminaire configurations, photometric files, technical metadata, and BIM models from Zumtobel. Delivered as clean JSON, CSV, or Parquet to S3, BigQuery, or Snowflake on your cadence.
Structured, schema-consistent data across all major object types — delivered clean, typed, and ready to query.
Complete list of extractable fields for Luminaire Specifications objects from zumtobel.com. All fields typed and schema-versioned.
"article_number": "42187391", "product_family": "PANOS infinity", "light_source": "LED", "luminous_flux_lm": 2400, "connected_load_w": 22, "luminous_efficacy_lm_w": 109, "colour_temperature_k": 4000, "cri": 90
| # | article_number | product_family | description | light_source | luminous_flux_lm | connected_load_w |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Photometric & Asset Files objects from zumtobel.com. All fields typed and schema-versioned.
"article_number": "42187391", "ldt_url": "https://example.com/panos.ldt", "ies_url": "https://example.com/panos.ies", "revit_url": "https://example.com/panos.rfa", "datasheet_pdf_url": "https://example.com/ds.pdf", "image_urls": "['https://example.com/img1.jpg']"
| # | article_number | ldt_url | ies_url | revit_url | autocad_url | dialux_url |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Product Configurator objects from zumtobel.com. All fields typed and schema-versioned.
"base_article_number": "PANOS_INF", "variant_id": "PI_R150_24W_840", "mounting_type": "Recessed", "housing_colour": "White", "optics_type": "Reflector", "beam_angle": "Flood 60°", "control_gear": "DALI", "dimming_protocol": "DALI-2"
| # | base_article_number | variant_id | mounting_type | housing_colour | optics_type | beam_angle |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Eco-Design & Sustainability objects from zumtobel.com. All fields typed and schema-versioned.
"article_number": "42187391", "energy_efficiency_class": "C", "eprel_registration_number": "876342", "lifetime_l90_b50": 50000, "replaceable_light_source": true, "replaceable_control_gear": true, "recyclable_pct": 94.5
| # | article_number | energy_efficiency_class | eprel_registration_number | lifetime_l90_b50 | lifetime_l80_b50 | replaceable_light_source |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Categories & Taxonomy objects from zumtobel.com. All fields typed and schema-versioned.
"category_name": "Recessed luminaires", "parent_category": "Indoor lighting", "application_area": "Office & Communication", "product_line": "Downlights", "series": "PANOS", "sub_series": "PANOS infinity R150", "related_systems": "['LITECOM']"
| # | category_id | category_name | parent_category | application_area | product_line | series |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Our pipeline handles complex industrial catalogues: stateful product configurators, nested technical specifications, and bulk asset downloads for photometric files and BIM models.
Lumens, wattage, CRI, MacAdam ellipse, and IP ratings scraped at the precise article number level.
Batch download and URL mapping for LDT, IES, and EULUMDAT light distribution files.
Direct extraction of Revit families (.rfa), AutoCAD (.dwg), and 3D models linked to specific configurations.
Crawl dynamic product configurators to extract all valid combinations of optics, control gear, and mounting types.
Extract region-specific catalogues across zumtobel.com/gb-en, /de-de, /us-en, handling local compliance metrics.
Capture EPREL registration numbers, energy efficiency classes, and component replaceability data.
Map complex modular systems like TECTON trunking into hierarchical relational schemas.
Automated checks for dead links on PDFs, mounting instructions, and photometric downloads.
Track new product introductions, phase-outs, and specification revisions on a weekly or monthly cadence.
Brief in. Clean data out.
Provide product families, application areas, or target regions. We design the extraction schema together.
We configure Scrapy and Playwright crawlers to handle Zumtobel's configurator logic and asset download queues.
Schema validation, unit standardisation, and null-rate checks before full launch.
JSON, CSV, or Parquet pushed to your S3 bucket, BigQuery dataset, or Snowflake stage on agreed cadence.
Extracting data from B2B manufacturing sites involves heavy payloads and dynamic configurators. Here is how we optimise the process.
Downloading thousands of photometric and BIM files requires heavy bandwidth and connection pooling. We map URLs and manage batch downloads asynchronously via S3 streams to prevent pipeline bottlenecks.
Zumtobel's product configurator uses complex JavaScript state to validate combinations. We use Playwright to iterate valid states and extract exact article numbers for every legal configuration.
Technical terms vary across regional sites. Our pipeline maps localised fields into a unified, English-normalised schema to ensure consistent queries across regions.
Technical specifications are often buried in accordion structures and dynamically loaded tabs. We target the underlying JSON payloads where possible, falling back to DOM parsing.
Modular systems like TECTON have parent-child-accessory relationships. We build relational keys linking trunking rails, luminaires, and sensors into a coherent database structure.
Import structured photometric data directly into DIALux, Relux, or proprietary calculation tools.
Compare luminous efficacy, pricing tiers, and warranty terms across architectural lighting manufacturers.
Populate architectural and MEP engineering databases with accurate Revit models and dimensional data.
Wholesalers and lighting distributors automate product catalogue updates, capturing new article numbers and spec changes.
Audit energy efficiency classes, EPREL data, and material circularity for green building certifications.
Extract modular system components to build accurate bills of materials for commercial lighting tenders.
"Zumtobel's digital catalogue contains critical photometric data and BIM models for architectural planning, but accessing it systematically requires navigating complex configurators and heavy asset payloads."
Most teams underestimate the complexity of scraping industrial manufacturers: handling stateful product configurators, normalising technical units, and managing asynchronous downloads for thousands of CAD and LDT files. DataFlirt absorbs that complexity so your engineers can focus on integrating the data into your planning tools, not maintaining crawler infrastructure.
Everything supported by our zumtobel.com scraper — rendered SPA elements, auth walls, rate-limit evasion and beyond.
Open-source tooling on proven cloud infra — no vendor lock-in, full observability.
Scrapy handles crawl orchestration and asset queues. Playwright handles JavaScript rendering for product configurators and dynamic tabs.
Dedicated worker nodes handle high-bandwidth downloads for BIM models and photometric files, streaming directly to S3 sinks.
Pipelines run on AWS Lambda (burst) and ECS (sustained). Airflow handles scheduling, dependency management, and SLA alerting.
Data delivered to where your team already works — no new tooling required.
About zumtobel.com scraping, legality, and pipeline operations.
Ask us directly →Scraping publicly available product specifications and photometric files from Zumtobel is generally permissible. DataFlirt targets only public, non-authenticated technical data. We do not extract personal data or circumvent authentication walls like the myZumtobel portal.
Yes. We map the exact file URLs to the corresponding article numbers and can optionally download the binaries directly to your S3 bucket for immediate integration into DIALux or Relux.
We use Playwright to programmatically select valid combinations of optics, control gear, and mounting types, extracting the generated article number and specifications for each valid state.
Yes. We capture links to Revit families (.rfa) and 2D/3D AutoCAD files associated with the product families and specific article numbers.
Yes. We build relational schemas that link trunking rails, nodes, luminaires, and emergency components together, ensuring the hierarchy remains intact in your database.
For architectural lighting catalogues, we typically run weekly or monthly diffs to capture new product launches, energy efficiency class updates, and specification revisions.
20-minute scoping call. Pilot dataset within the week. Production within two. Whether you need a one-off extraction of photometric files or a continuous feed of luminaire specifications, we scope, build, and operate the pipeline. Tell us what you need.