SYSTEM all green source motion.com queue 14,892 categories p99 latency 218ms dataflirt.com · scraper/motion-com
RUN · 42 active pipelines · motion.com live

Industrial parts data,
at warehouse scale.

We extract bearings, pneumatics, hydraulics, and power transmission catalogues from Motion.com. Delivered as clean JSON, CSV, or Parquet to S3, BigQuery, or Snowflake on your cadence.

SKUs extracted
4.2M /run
Inventory updates
8.1M /24h
Datasheets parsed
612K /week
Active pipelines
42
Uptime
99.94%
Data Dictionary

Every field we extract from motion.com

Structured, schema-consistent data across all major object types — delivered clean, typed, and ready to query.

Complete list of extractable fields for Product Core objects from motion.com. All fields typed and schema-versioned.

skumpnbrandtitledescriptioncategory_pathlist_pricecurrencyuommin_order_qty
product_core
● 200 OK
"sku": "00123456",
"mpn": "6204-2RS",
"brand": "SKF",
"title": "Deep Groove Ball Bearing",
"list_price": 14.5,
"currency": "USD",
"uom": "EA",
"min_order_qty": 1
# skumpnbrandtitledescriptioncategory_path
1
2
3

Complete list of extractable fields for Technical Specs objects from motion.com. All fields typed and schema-versioned.

skuweight_lbsmaterialbore_diameteroutside_diameterwidthstatic_load_capacitydynamic_load_capacitymax_rpmoperating_temp
technical_specs
● 200 OK
"sku": "00123456",
"bore_diameter": "20 mm",
"outside_diameter": "47 mm",
"width": "14 mm",
"material": "Steel",
"max_rpm": "10000",
"weight_lbs": 0.23
# skuweight_lbsmaterialbore_diameteroutside_diameterwidth
1
2
3

Complete list of extractable fields for Inventory & Branch objects from motion.com. All fields typed and schema-versioned.

skubranch_idbranch_namezip_codein_stockstock_qtylead_time_daysdrop_ship_eligibleshipping_weight
inventory_& branch
● 200 OK
"sku": "00123456",
"branch_id": "BR-104",
"branch_name": "Chicago North",
"zip_code": "60601",
"in_stock": true,
"stock_qty": 45,
"lead_time_days": 1
# skubranch_idbranch_namezip_codein_stockstock_qty
1
2
3

Complete list of extractable fields for Media & Documents objects from motion.com. All fields typed and schema-versioned.

skuprimary_image_urlgallery_imagesdatasheet_urlcad_2d_urlcad_3d_urlsafety_data_sheet_urlmanual_url
media_& documents
● 200 OK
"sku": "00123456",
"primary_image_url": "https://motion.com/img/00123456_main.jpg",
"datasheet_url": "https://motion.com/docs/SKF_6204_specs.pdf",
"cad_2d_url": "https://motion.com/cad/00123456_2d.dwg",
"cad_3d_url": "https://motion.com/cad/00123456_3d.step",
"safety_data_sheet_url": "None",
"manual_url": "None"
# skuprimary_image_urlgallery_imagesdatasheet_urlcad_2d_urlcad_3d_url
1
2
3

Complete list of extractable fields for Taxonomy & Cross Ref objects from motion.com. All fields typed and schema-versioned.

skuprimary_categorysub_categoryfamilyunspsc_codereplacement_skualternative_skusrelated_parts
taxonomy_& cross ref
● 200 OK
"sku": "00123456",
"primary_category": "Bearings",
"sub_category": "Ball Bearings",
"family": "Deep Groove",
"unspsc_code": "31171504",
"replacement_sku": "00123457",
"alternative_skus": "['00987654', '00554433']"
# skuprimary_categorysub_categoryfamilyunspsc_codereplacement_sku
1
2
3

Capabilities

Industrial part data extracted at scale

Our Motion scraper navigates complex category trees, parses variable technical specifications, and captures localised inventory data across thousands of branches.

Full SKU Extraction

Extract Motion item numbers, manufacturer part numbers, UPCs, and brand identifiers for every component.

Deep Taxonomy Crawling

Traverse the massive MRO category hierarchy to map every part to its exact family and subcategory.

Technical Specification Parsing

Normalise variable spec tables into structured JSON. Capture bore sizes, load capacities, and operating temperatures.

Branch Inventory Localisation

Simulate location states to extract accurate stock depth and availability for specific postal codes and branches.

Document & CAD Links

Capture direct URLs for PDF datasheets, safety data sheets, and 2D/3D CAD models linked to each SKU.

List Pricing & UOM

Extract standard list prices, currency, minimum order quantities, and units of measure.

Cross Reference Mapping

Scrape alternative parts, direct replacements, and related components to build comprehensive cross reference databases.

Brand Aggregation

Filter and extract catalogues for specific manufacturers across the entire Motion distribution network.

Incremental Updates

Run daily diffs against millions of SKUs to capture price changes and inventory fluctuations without full re-crawls.

// engagement pipeline

From category list to warehouse record

Brief in. Clean data out.

Define Scope
d 0

Provide target categories, brands, or MPN lists. We design the extraction schema together.

Pipeline Build
d 2–4

We configure Scrapy crawlers, proxy rotation, and session management for motion.com.

Validation & QA
d 4–6

Schema validation, null rate checks, and spec normalisation testing before full launch.

Delivery
ongoing

JSON, CSV, or Parquet pushed to your S3 bucket, BigQuery dataset, or Snowflake stage on agreed cadence.

Under the hood

How our Motion pipeline handles the hard parts

Industrial distributors use strict bot mitigation and complex localised states. Here is how we maintain reliable extraction.

pipeline-monitor · motion.com · live ● active
// fingerprinting
Identity rotation
TLS fingerprintrandomised
User-agentrotated
IP poolresidential
Challenges blocked0
// pagination
Page coverage
48,291 pages queued running
// observability
Pipeline health
99.9%
uptime
142ms
p99 lat
0.3%
null rate
2
alerts
Anti-bot layer
Residential proxy rotation

Motion employs strict rate limiting and bot detection. Our crawlers use US residential ISP proxies with realistic browser fingerprints and randomised request timing to maintain high success rates.

Localised state
Branch specific session handling

Inventory and pricing vary by location. We manage HTTP cookie sessions to simulate specific postal codes, ensuring you receive accurate branch level stock data.

Taxonomy navigation
Deep pagination handling

MRO catalogues feature thousands of subcategories and deep pagination. Our orchestrator maps the complete taxonomy tree before crawling to ensure zero missed SKUs.

Unstructured data
Specification table normalisation

Technical specifications vary wildly between bearings and pneumatics. We use custom parsing logic to normalise diverse HTML tables into consistent, strongly typed JSON fields.

Change detection
Hash based diffing

For catalogues exceeding 4M SKUs, we maintain a hash index of last seen values. Subsequent runs only push diffs, reducing your downstream processing load.

Applications

Who uses Motion data

Teams across industries use motion.com data to build competitive products and smarter operations.

01
Competitor Price Monitoring

Distributors track Motion list prices across key brands to optimise their own pricing strategies.

02
Cross Reference Database Building

Manufacturers map their MPNs against Motion alternatives to improve their internal cross reference tools.

03
Procurement Optimisation

Supply chain teams monitor branch level inventory to identify reliable secondary sources for critical components.

04
Assortment Gap Analysis

Market researchers analyse category depth to identify whitespace and new product introduction opportunities.

05
ERP & PIM Enrichment

B2B merchants enrich their Product Information Management systems with normalised specifications and datasheet URLs.

06
AI & Search Index Training

Data science teams train internal search algorithms on structured MRO taxonomy and technical specifications.

Why DataFlirt

"Motion.com holds the definitive taxonomy for industrial parts, but mapping millions of MPNs and specifications requires purpose-built extraction infrastructure."

Extracting MRO data at scale means navigating millions of SKUs, complex category trees, and branch specific inventory states. DataFlirt handles the proxy rotation, session management, and schema normalisation so your engineers receive clean, warehouse ready part data.

Technical Spec

Motion scraper technical capabilities

Everything supported by our motion.com scraper — rendered SPA elements, auth walls, rate-limit evasion and beyond.

JavaScript rendering
Playwright sessions required for dynamic inventory widgets
Supported
Residential proxy rotation
US residential IPs rotated per request
Supported
Branch level inventory
Stock depth captured per specified postal code
Supported
Specification normalisation
Variable tables mapped to typed JSON keys
Supported
Datasheet extraction
Direct URLs to PDFs and CAD models
Supported
Cross reference mapping
Alternative and replacement SKUs captured
Supported
Change detection
Hash based diffing for incremental updates
Supported
Contract pricing
Account specific negotiated pricing requires customer credentials
Partial
B2B Cart checkout simulation
Automated order placement is strictly prohibited
Partial
Infrastructure

Infrastructure powering the Motion pipeline

Open-source tooling on proven cloud infra — no vendor lock-in, full observability.

ScrapyPlaywrightPython 3.12RedisPostgreSQLApache AirflowAWS LambdaS3CloudWatch2CaptchaCapSolverResidential ProxiesDockerKubernetesGrafanaPrometheus
Scrapy + Playwright Stack

Scrapy handles crawl orchestration and deduplication. Playwright handles JavaScript rendering and location specific cookie sessions.

Residential Proxy Infrastructure

We maintain pools of US residential ISP proxies. Rotation happens per request with sticky sessions for branch inventory extraction.

Cloud Native Orchestration

Pipelines run on AWS Lambda and ECS. Airflow handles scheduling and dependency management. All state stored in managed Postgres.

Output & Delivery

Your data, your destination

Data delivered to where your team already works — no new tooling required.

JSON
Newline delimited or nested schema versioned per run
CSV
Flat file with typed columns Excel compatible
XLS
Standard spreadsheet format for business users
Parquet
Columnar format for BigQuery and Snowflake
AWS S3
Direct bucket delivery compatible with any data lake
Webhook
HTTP POST per record for real time processing
API
REST endpoints to query extracted datasets
BigQuery
Streamed directly into your dataset
Snowflake
Stage and COPY INTO workflow
Postgres
Upsert into your existing schema
S3
Direct bucket delivery — compatible with any data lake
// faq

Common questions.

About motion.com scraping, legality, and pipeline operations.

Ask us directly →
Is scraping Motion legal?

Scraping publicly available catalogue information is generally permissible under applicable law. DataFlirt targets only public, non authenticated product, pricing, and specification data. We do not extract personal data or circumvent authentication walls.

How do you handle bot detection on Motion?

We use US residential proxies, realistic browser fingerprints, and request timing modelled on human behaviour. We monitor for block rate spikes in real time and trigger pool rotation automatically.

Can you extract inventory for specific branches?

Yes. We configure our crawlers to simulate specific postal codes, allowing us to capture accurate stock availability and lead times for your target locations.

How fresh is the data?

Full catalogue refreshes typically complete within a 24 to 48 hour window depending on scale. Targeted category pipelines can run at hourly cadences.

Do you standardise technical specifications?

Yes. We map variable HTML specification tables into consistent, strongly typed JSON fields, ensuring bore diameters and load capacities are uniformly formatted.

What is the minimum viable engagement?

Our smallest packages start at a defined category list or MPN set with weekly delivery. For full site extraction, we price based on volume and delivery frequency.

Can I request a sample dataset?

Absolutely. We provide a sample run of up to 500 SKUs as part of the scoping process so you can validate schema fit and data quality.

$ dataflirt scope --new-project --source=motion.com ready

Tell us what
to extract.
We do the rest.

20-minute scoping call. Pilot dataset within the week. Production within two. Whether you need a specific category extraction or a continuous feed across 4M MRO parts, we scope, build, and operate the pipeline. Tell us what you need.

hello@dataflirt.com · Bengaluru · IST · typical reply < 4h
Related Scrapers

More in industrial and mro

Services

Data Extraction for Every Industry

View All Services →