We extract SKUs, sizing availability, pricing signals, and material data from only.com. Delivered as clean JSON, CSV, or Parquet to S3, BigQuery, or Snowflake on your cadence.
Structured, schema-consistent data across all major object types — delivered clean, typed, and ready to query.
Complete list of extractable fields for Product Listings objects from only.com. All fields typed and schema-versioned.
"sku": "15283940", "title": "ONLPOPTRASH EASY PANT", "brand": "ONLY", "category": "Trousers", "price": 39.99, "currency": "EUR", "colour": "Black", "sizes": "['XS', 'S', 'M', 'L', 'XL']"
| # | sku | title | brand | category | price | currency |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Inventory & Sizing objects from only.com. All fields typed and schema-versioned.
"sku": "15283940", "size": "M", "in_stock": true, "stock_quantity": 14, "low_stock_warning": false, "expected_restock": "None", "variant_id": "15283940-M-BLK"
| # | sku | size | in_stock | stock_quantity | low_stock_warning | expected_restock |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Pricing & Promotions objects from only.com. All fields typed and schema-versioned.
"sku": "15283940", "base_price": 49.99, "current_price": 39.99, "discount_pct": 20, "discount_abs": 10.0, "promotion_name": "Mid Season Sale", "currency": "EUR", "region": "DE"
| # | sku | base_price | current_price | discount_pct | discount_abs | promotion_name |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Materials & Care objects from only.com. All fields typed and schema-versioned.
"sku": "15283940", "primary_material": "63% Viscose", "secondary_material": "32% Nylon, 5% Elastane", "sustainability_label": "Livaeco by Birla Cellulose", "washing_temp": "30°C mild fine wash", "ironing_rules": "Iron at moderate temperature", "dry_clean": false
| # | sku | primary_material | secondary_material | sustainability_label | washing_temp | ironing_rules |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Category & Hierarchy objects from only.com. All fields typed and schema-versioned.
"sku": "15283940", "breadcrumbs": "['Home', 'Clothing', 'Trousers', 'Casual Trousers']", "primary_category": "Clothing", "sub_category": "Trousers", "collection_name": "Poptrash", "season": "AW23", "gender": "Female"
| # | sku | breadcrumbs | primary_category | sub_category | collection_name | season |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Our pipeline handles the dynamic frontend of only.com, extracting complete SKU matrices, geo-specific pricing, and real-time inventory levels without getting blocked by regional redirects.
Title, description, materials, care instructions, and fit details extracted directly from the product detail pages.
Map parent products to child SKUs across every colour and size combination, maintaining the multidimensional matrix.
Monitor size-level availability and low-stock indicators to track sell-through rates and inventory depth.
Extract accurate pricing across different European and global regions by routing traffic through local ISP proxies.
Capture exact fabric compositions and sustainability labels (e.g., Livaeco) for compliance and ESG reporting.
Extract clean URLs for all product imagery, including front, back, detail, and model shots.
Track markdown events, discount percentages, and promotional badges applied to specific SKUs.
Extract full breadcrumb trails and collection names to maintain accurate taxonomy mapping.
Run extractions daily or hourly to catch fast-moving fashion trends and flash sales.
Brief in. Clean data out.
Provide category URLs, specific collections, or target regions. We design the extraction schema together.
We configure Scrapy / Playwright crawlers, proxy rotation, and session management for only.com.
Schema validation, null-rate checks, and variant mapping verification before full launch.
JSON / CSV / Parquet pushed to your S3 bucket, BigQuery dataset, or Snowflake stage on agreed cadence.
Fashion eCommerce relies heavily on dynamic rendering and geo-blocking. Here is how we maintain stable extraction.
Only.com aggressively redirects users based on IP location to show regional pricing and stock. We use strict country-targeted residential proxies to force the crawler into the correct locale, ensuring you get accurate EU, UK, or global data.
Fashion SKUs are nested. A single product URL can contain dozens of colour and size combinations, each with distinct stock levels. We execute the necessary JavaScript state changes to extract the full matrix without missing hidden variants.
Category pages use infinite scroll and dynamic loading. Instead of brittle DOM scrolling, we intercept the underlying XHR requests to paginate cleanly through the entire catalogue, ensuring zero dropped items.
We maintain a hash index of last-seen values per SKU. Subsequent runs only push diffs — highly efficient for tracking daily price markdowns and stock changes without processing the entire catalogue.
Bestseller frequently updates their frontend architecture. We monitor for null-rate spikes in critical fields like price and stock, automatically alerting our engineers to update selectors before you receive malformed data.
Fashion retailers track Only's pricing strategy, seasonal markdowns, and promotional events to optimise their own pricing.
Analysts monitor size-level stock depletion to identify fast-selling items and predict upcoming fashion trends.
Merchandisers analyse category depth, colour availability, and sizing curves to plan competing assortments.
Pricing teams correlate stock depth with discount percentages to understand Bestseller's clearance strategies.
ESG analysts track the adoption of sustainable materials (like Livaeco) across the catalogue over time.
Computer vision teams use structured high-res image datasets mapped to specific clothing categories and colours for model training.
"Fast fashion moves quickly. Stock depth and pricing on only.com fluctuate daily across regions, requiring persistent, high-frequency extraction to capture market realities."
Extracting data from Bestseller brands like Only requires navigating complex JavaScript frontends, aggressive geo-redirects, and multidimensional SKU matrices. DataFlirt handles the proxy routing and variant normalisation so your analysts get clean warehouse records, not broken scripts.
Everything supported by our only.com scraper — rendered SPA elements, auth walls, rate-limit evasion and beyond.
Open-source tooling on proven cloud infra — no vendor lock-in, full observability.
Scrapy handles crawl orchestration and deduplication. Playwright handles JavaScript rendering, infinite scroll, and variant selection.
We maintain pools of residential ISP proxies mapped to specific European and global regions to bypass Only's IP-based redirects.
Pipelines run on AWS Lambda and ECS. Airflow handles scheduling and dependency management. All state stored in managed Postgres.
Data delivered to where your team already works — no new tooling required.
About only.com scraping, legality, and pipeline operations.
Ask us directly →Scraping publicly available information from only.com is generally permissible. DataFlirt targets only public, non-authenticated product, pricing, and inventory data. We do not extract personal data, circumvent authentication walls, or violate GDPR.
Only.com redirects traffic based on IP origin. We use country-specific residential proxies to ensure our crawlers access the correct regional storefront, capturing accurate local pricing and currency.
Yes. Our pipeline maps the entire SKU matrix. We extract data for every available size and colour combination linked to a parent product, including individual stock statuses.
We can configure pipelines to run at daily, sub-daily, or hourly cadences depending on your requirements. High-frequency runs are typically scoped to specific high-value categories to monitor rapid stock depletion.
Yes. We extract primary and secondary material compositions, along with care instructions and sustainability tags (e.g., Livaeco, recycled materials) directly from the product details.
Our packages start at defined category or collection tracking with weekly delivery. For full-catalogue daily refreshes, we price based on compute volume and delivery frequency.
Yes. We provide a sample run of up to 500 SKUs as part of the pre-engagement scoping process, allowing you to validate schema fit and data quality.
20-minute scoping call. Pilot dataset within the week. Production within two. Whether you need a one-off catalogue dump or continuous price and inventory monitoring — we scope, build, and operate the pipeline. Tell us what you need.