SYSTEM all green source otto.de queue 18,492 pages p99 latency 218ms dataflirt.com · scraper/otto-de
RUN / 64 active pipelines / otto.de live

Otto.de data,
at warehouse scale.

We extract product listings, variant availability, marketplace seller pricing, and customer reviews from Otto.de. Delivered as clean JSON, CSV, or Parquet to S3 or BigQuery on your cadence.

Products extracted
840K /day
Price updates
2.1M /24h
Review records
315K /run
Active pipelines
64
Uptime
99.94%
Data Dictionary

Every field we extract from otto.de

Structured, schema-consistent data across all major object types — delivered clean, typed, and ready to query.

Complete list of extractable fields for Product Listings objects from otto.de. All fields typed and schema-versioned.

skutitlebrandcategorysub_categorybase_pricecurrent_pricecurrencydiscount_pctcoloursizematerialeanin_stockratingreview_countdescriptionenergy_rating
product_listings
● 200 OK
"sku": "84920184A",
"title": "Adidas Originals Sneaker",
"brand": "Adidas",
"current_price": 89.99,
"currency": "EUR",
"discount_pct": 10,
"colour": "White",
"in_stock": true
# skutitlebrandcategorysub_categorybase_price
1
2
3

Complete list of extractable fields for Pricing & Variants objects from otto.de. All fields typed and schema-versioned.

skuvariant_idcoloursizebase_pricecurrent_pricediscount_pctdelivery_timeseller_nameshipping_coststock_status
pricing_& variants
● 200 OK
"sku": "84920184A",
"variant_id": "V918237",
"colour": "White",
"size": "42",
"current_price": 89.99,
"delivery_time": "2 to 3 workdays",
"seller_name": "Otto",
"shipping_cost": 2.95
# skuvariant_idcoloursizebase_pricecurrent_price
1
2
3

Complete list of extractable fields for Reviews & Ratings objects from otto.de. All fields typed and schema-versioned.

review_idskuratingtitletextdateverified_purchasehelpful_votesauthor_namefit_feedback
reviews_& ratings
● 200 OK
"review_id": "REV99281",
"sku": "84920184A",
"rating": 5,
"title": "Great fit and quality",
"text": "The sneakers fit perfectly and look exactly like the pictures.",
"date": "2026-03-14",
"verified_purchase": true,
"helpful_votes": 12
# review_idskuratingtitletextdate
1
2
3

Complete list of extractable fields for Seller Intelligence objects from otto.de. All fields typed and schema-versioned.

seller_idseller_nameratingreview_countproducts_countreturn_policylegal_nameshipping_methodsbusiness_addresscontact_email
seller_intelligence
● 200 OK
"seller_id": "SEL4492",
"seller_name": "SneakerWorld GmbH",
"rating": 4.8,
"review_count": 1420,
"products_count": 350,
"return_policy": "30 days free returns",
"legal_name": "SneakerWorld Retail GmbH",
"shipping_methods": "Hermes, DHL"
# seller_idseller_nameratingreview_countproducts_countreturn_policy
1
2
3

Complete list of extractable fields for Search Results objects from otto.de. All fields typed and schema-versioned.

keywordpositionskutitlepricebrandsponsoredratingreview_countimage_url
search_results
● 200 OK
"keyword": "white sneakers",
"position": 3,
"sku": "84920184A",
"sponsored": false,
"title": "Adidas Originals Sneaker",
"price": 89.99,
"rating": 4.6,
"review_count": 312
# keywordpositionskutitlepricebrand
1
2
3

Capabilities

Complete Otto.de marketplace coverage

Our infrastructure extracts the full Otto.de catalogue: fashion variants, home appliance specifications, marketplace seller offers, and dynamic pricing.

Full Catalogue Extraction

Title, brand, category, description, specifications, and images extracted across fashion, living, and electronics categories.

Variant Mapping

Extract all colour and size combinations per product, mapping base SKUs to their respective child variants.

Marketplace Seller Tracking

Identify third-party sellers on Otto.de, capturing seller ratings, review counts, and shipping policies.

Dynamic Pricing & Discounts

Track base price, current price, discount percentages, and shipping costs timestamped per crawl.

Delivery & Stock Status

Capture estimated delivery windows, stock availability, and specific shipping carrier information.

Review Mining

Extract customer ratings, review text, verified purchase flags, and specific fit feedback for apparel.

Search Rank Tracking

Monitor keyword positions, distinguishing between organic results and sponsored placements.

Energy Ratings & Specs

Extract structured technical specifications and EU energy efficiency classes for home appliances.

Scheduled Pipelines

Run extractions at daily or weekly intervals with change detection to capture pricing shifts.

// engagement pipeline

From URL list to warehouse record

Brief in. Clean data out.

Define Scope
d 0

Provide category URLs, brand names, or keyword sets. We design the extraction schema together.

Pipeline Build
d 2–4

We configure Scrapy and Playwright crawlers, proxy rotation, and session management for Otto.de.

Validation & QA
d 4–6

Schema validation, null-rate checks, and data normalisation before full launch.

Delivery
ongoing

JSON, CSV, or Parquet pushed to your S3 bucket or data warehouse on agreed cadence.

Under the hood

Handling Otto.de bot protection

Otto.de employs strict rate limiting and fingerprinting. We handle the infrastructure complexity so you receive clean data.

pipeline-monitor · otto.de · live ● active
// fingerprinting
Identity rotation
TLS fingerprintrandomised
User-agentrotated
IP poolresidential
Challenges blocked0
// pagination
Page coverage
48,291 pages queued running
// observability
Pipeline health
99.9%
uptime
142ms
p99 lat
0.3%
null rate
2
alerts
Anti-bot layer
Residential proxy rotation

We route requests through German residential IPs to match expected geographic behaviour, preventing IP blocks and rate limits.

JavaScript rendering
Playwright execution

Otto.de relies on client-side rendering for variant pricing and availability. We use headless browsers to execute JavaScript and capture accurate DOM states.

Data completeness
Variant hydration

We iterate through size and colour selectors to expose specific pricing and stock levels that are hidden in the initial page load.

Schema stability
Resilient selectors

Our extraction logic uses multiple fallback chains per field, ensuring data flows even when Otto.de updates its frontend framework.

Monitoring
Anomaly detection

Every run emits structured logs. We monitor for null-rate spikes and schema drift, addressing issues before they impact your warehouse.

Applications

Who uses Otto.de data

Teams across industries use otto.de data to build competitive products and smarter operations.

01
Price Intelligence

Retailers monitor Otto.de pricing and discount strategies to adjust their own marketplace positioning.

02
Brand Monitoring

Brands track how their products are presented, priced, and reviewed by third-party sellers on the Otto marketplace.

03
Assortment Planning

Merchandising teams analyse category depth, variant availability, and out-of-stock rates to identify market gaps.

04
Seller Analysis

Aggregators evaluate third-party seller performance, tracking rating velocity and catalogue size.

05
Review Sentiment

Product teams extract customer reviews to understand fit issues, material quality, and overall sentiment.

06
Search Optimisation

Agencies track keyword rankings and sponsored placements to optimise visibility for their brand clients.

Why DataFlirt

"Otto.de represents the core of German e-commerce. Extracting accurate variant and marketplace data requires continuous infrastructure maintenance."

Building an internal scraper for Otto.de means fighting bot protection and complex DOM structures. DataFlirt provides a managed extraction pipeline with residential proxies and headless browsers. Your engineering team gets structured warehouse data without the operational overhead.

Technical Spec

Otto.de scraper technical specifications

Everything supported by our otto.de scraper — rendered SPA elements, auth walls, rate-limit evasion and beyond.

JavaScript rendering
Playwright sessions required for variant pricing and dynamic stock status
Supported
CAPTCHA bypass
Automated solver integration for rate-limit challenges
Supported
Residential proxy rotation
German ISP proxies to maintain high success rates
Supported
Variant mapping
Extraction of all colour and size permutations per base product
Supported
Delivery estimates
Capture of specific shipping windows and carrier data
Supported
Seller mapping
Attribution of offers to specific marketplace sellers
Supported
Webhook delivery
HTTP POST per record for immediate downstream processing
Supported
User purchase history
Account-gated order data requires user credentials
Partial
Otto UP subscription details
Premium delivery subscription benefits tied to authenticated accounts
Partial
Infrastructure

Infrastructure powering the Otto.de pipeline

Open-source tooling on proven cloud infra — no vendor lock-in, full observability.

ScrapyPlaywrightPython 3.12RedisPostgreSQLApache AirflowAWS LambdaS3CloudWatch2CaptchaCapSolverResidential ProxiesDockerKubernetesGrafanaPrometheus
Scrapy + Playwright Stack

Scrapy manages crawl orchestration and retry logic. Playwright handles JavaScript execution and interaction flows required for variant hydration.

Residential Proxy Infrastructure

We maintain pools of German residential ISP proxies. Rotation happens per request to prevent IP bans and ensure consistent access to Otto.de.

Cloud-Native Orchestration

Pipelines run on AWS Lambda and ECS. Airflow handles scheduling and dependency management, ensuring data is delivered on your precise cadence.

Output & Delivery

Your data, your destination

Data delivered to where your team already works — no new tooling required.

JSON
Newline-delimited or nested arrays
CSV
Flat file with typed columns
XLS
Excel compatible format for business teams
Parquet
Columnar format for data warehouses
AWS S3
Direct bucket delivery
Webhook
HTTP POST per record
API
REST endpoint for on-demand querying
BigQuery
Streamed directly into your dataset
S3
Direct bucket delivery — compatible with any data lake
// faq

Common questions.

About otto.de scraping, legality, and pipeline operations.

Ask us directly →
Can you extract all variants for a specific fashion item?

Yes. We iterate through the colour and size selectors on Otto.de product pages to capture the specific price, stock status, and delivery time for every variant combination.

Do you capture third-party marketplace seller data?

Yes. Otto.de operates as a marketplace. We extract the seller name, seller rating, and specific shipping policies for every offer on a product page.

How do you handle Otto.de bot protection?

We use German residential proxies and headless Playwright browsers to mimic normal user behaviour, avoiding the rate limits that block standard datacenter IPs.

Can I get daily price updates?

Yes. We configure pipelines to run at daily intervals, using change detection to only push records where pricing or availability has shifted.

Do you extract customer reviews?

Yes. We paginate through the review sections to capture star ratings, text, date, and verified purchase indicators.

What formats do you deliver?

We deliver data in JSON, CSV, XLS, and Parquet. We can push directly to AWS S3, BigQuery, Snowflake, or trigger webhooks for immediate processing.

$ dataflirt scope --new-project --source=otto.de ready

Tell us what
to extract.
We do the rest.

20-minute scoping call. Pilot dataset within the week. Production within two. Whether you need a category extraction or daily price monitoring across thousands of variants, we scope, build, and operate the pipeline. Tell us what you need.

hello@dataflirt.com · Bengaluru · IST · typical reply < 4h
Related Scrapers

More in fashion and apparel

Services

Data Extraction for Every Industry

View All Services →