SYSTEM all green source netto-online.de queue 14,892 pages p99 latency 184ms dataflirt.com · scraper/netto-online-de
RUN · 31 active pipelines · netto-online.de live

Netto grocery data,
at warehouse scale.

We extract product listings, regional pricing, nutritional profiles, allergen warnings, and weekly promotional offers from Netto. Delivered as clean JSON, CSV, or Parquet to S3, BigQuery, or Snowflake on your cadence.

Products extracted
142K /day
Price updates
314K /24h
Weekly specials
12K /run
Active pipelines
31
Uptime
99.94%
Data Dictionary

Every field we extract from netto-online.de

Structured, schema-consistent data across all major object types — delivered clean, typed, and ready to query.

Complete list of extractable fields for Product Listings objects from netto-online.de. All fields typed and schema-versioned.

skutitlebrandcategorysub_categorypricebase_pricepfand_valuein_stockdescriptioningredientsallergensimage_urlsproduct_url
product_listings
● 200 OK
"sku": "218493000",
"title": "Coca-Cola Original Taste 1,25 Liter",
"brand": "Coca-Cola",
"price": 1.29,
"pfand_value": 0.25,
"in_stock": true,
"category": "Getraenke"
# skutitlebrandcategorysub_categoryprice
1
2
3

Complete list of extractable fields for Pricing & Promotions objects from netto-online.de. All fields typed and schema-versioned.

skucurrent_priceoriginal_pricediscount_pctpromotion_typeis_weekly_specialvalid_fromvalid_tobase_price_per_unitcurrency
pricing_& promotions
● 200 OK
"sku": "218493000",
"current_price": 1.29,
"original_price": 1.79,
"discount_pct": 27,
"is_weekly_special": true,
"base_price_per_unit": "1.03 EUR / Liter",
"valid_to": "2026-05-18T23:59:59Z"
# skucurrent_priceoriginal_pricediscount_pctpromotion_typeis_weekly_special
1
2
3

Complete list of extractable fields for Nutritional Info objects from netto-online.de. All fields typed and schema-versioned.

skuenergy_kjenergy_kcalfat_gsaturated_fat_gcarbs_gsugar_gprotein_gsalt_gfibre_g
nutritional_info
● 200 OK
"sku": "218493000",
"energy_kj": 180,
"energy_kcal": 42,
"fat_g": 0.0,
"carbs_g": 10.6,
"sugar_g": 10.6,
"protein_g": 0.0,
"salt_g": 0.0
# skuenergy_kjenergy_kcalfat_gsaturated_fat_gcarbs_g
1
2
3

Complete list of extractable fields for Category Hierarchy objects from netto-online.de. All fields typed and schema-versioned.

category_idcategory_nameparent_categorylevelproduct_counturlbreadcrumbactive_promotions
category_hierarchy
● 200 OK
"category_id": "cat_10294",
"category_name": "Cola & Limonade",
"parent_category": "Alkoholfreie Getraenke",
"level": 3,
"product_count": 142,
"breadcrumb": "Startseite > Getraenke > Alkoholfreie Getraenke > Cola & Limonade"
# category_idcategory_nameparent_categorylevelproduct_counturl
1
2
3

Complete list of extractable fields for Regional Availability objects from netto-online.de. All fields typed and schema-versioned.

skustore_idzip_codecityavailablestock_levellocal_pricedelivery_time
regional_availability
● 200 OK
"sku": "218493000",
"zip_code": "10115",
"city": "Berlin",
"available": true,
"stock_level": "HIGH",
"local_price": 1.29,
"delivery_time": "1-3 Werktage"
# skustore_idzip_codecityavailablestock_level
1
2
3

Capabilities

Everything you need from Netto, nothing you don't

Our Netto scraper handles every layer of the platform: grocery listings, dynamic regional pricing, nutritional profiles, and weekly promotional cycles, with JavaScript rendering and bot circumvention built in.

Full Product Extraction

Extract SKU, title, descriptions, imagery, and private label identification across the entire Netto catalogue.

Regional Price Tracking

Capture local pricing variations and availability based on specific PLZ/zip code injections.

Promotional Cycles

Track weekly 'Aktionen' and discount windows with precise start and end timestamps.

Nutritional & Allergen Data

Extract structured tables for macronutrients, ingredients lists, and mandatory allergen warnings.

Pfand Calculation

Isolate mandatory bottle deposits from the base price for accurate total cost calculation.

Base Price Normalisation

Extract price per kilogram or litre to enable accurate cross-brand and cross-retailer comparison.

Stock & Availability

Monitor warehouse inventory status, stock depth indicators, and regional delivery estimates.

Category Mapping

Reconstruct Netto's full taxonomy from root nodes down to leaf categories and product counts.

Scheduled Modes

Run continuous pipelines aligned with Netto's weekly circular updates and daily price adjustments.

// engagement pipeline

From SKU list to warehouse record

Brief in. Clean data out.

Define Scope
d 0

Provide category URLs, zip codes, or SKU lists. We design the extraction schema together.

Pipeline Build
d 2–4

We configure Scrapy / Playwright crawlers, proxy rotation, and session management for netto-online.de.

Validation & QA
d 4–6

Schema validation, null-rate checks, and price-outlier detection before full launch.

Delivery
ongoing

JSON / CSV / Parquet pushed to your S3 bucket, BigQuery dataset, or Snowflake stage on agreed cadence.

Under the hood

How our Netto pipeline handles the hard parts

Grocery retail sites use strict caching and bot protection to defend their pricing data. Here is how we stay resilient and maintain uptime.

pipeline-monitor · netto-online.de · live ● active
// fingerprinting
Identity rotation
TLS fingerprintrandomised
User-agentrotated
IP poolresidential
Challenges blocked0
// pagination
Page coverage
48,291 pages queued running
// observability
Pipeline health
99.9%
uptime
142ms
p99 lat
0.3%
null rate
2
alerts
Geo-targeted proxies
German residential IPs for accurate pricing

Routing requests through German residential IPs ensures we access accurate regional pricing, avoid geo-blocking, and bypass basic data centre IP bans.

Session state management
Zip code injection and cookie handling

Netto alters pricing and stock based on location. We manage cookie consent and inject target zip codes into the session state to render localised product variants.

JavaScript rendering
Playwright execution for dynamic stock

We execute full Playwright sessions to hydrate dynamic promotional widgets, stock indicators, and lazy-loaded nutritional tables that simple HTTP clients miss.

Schema stability
Fallback chains for layout shifts

The weekly specials section frequently changes DOM structure. We maintain fallback selectors for pricing and product details to prevent pipeline breakages.

Change detection
Hash-based diffing for daily sweeps

We hash nutritional and pricing fields to emit only modified records during daily catalog sweeps, reducing downstream processing load.

Applications

Who uses Netto data and how

Teams across industries use netto-online.de data to build competitive products and smarter operations.

01
Competitor Price Matching

Supermarkets track Netto's private label pricing and promotional calendars to adjust their own weekly circulars.

02
FMCG Brand Monitoring

Consumer goods brands audit shelf presence, discount frequency, and category share across Netto's digital storefront.

03
Inflation Tracking

Economic analysts monitor basket price indices across essential grocery categories to measure consumer inflation trends.

04
Nutritional Analysis

Health tech companies ingest macronutrient and allergen data to power dietary recommendation engines and food-scoring apps.

05
Supply Chain Forecasting

Logistics providers correlate stockout indicators with regional demand signals to optimise distribution routes.

06
Retail Strategy

Consultancies analyse Netto's promotional depth versus standard pricing to benchmark discount retail models in the DACH region.

Why DataFlirt

"Netto-Online.De holds critical signals for German retail pricing and FMCG brand visibility, but extracting it requires navigating strict regional session states."

Most teams underestimate the investment required: reliable grocery scraping requires German residential proxies, full JavaScript rendering for zip code injection, and daily selector maintenance. DataFlirt absorbs that complexity so your engineers can focus on the analysis, not the infrastructure.

Technical Spec

Netto scraper technical capabilities

Everything supported by our netto-online.de scraper — rendered SPA elements, auth walls, rate-limit evasion and beyond.

JavaScript rendering
Playwright sessions for dynamic stock and price hydration
Supported
German residential proxies
ISP-grade IPs from DE pools to prevent blocking
Supported
Zip code injection
Localised pricing via cookie and header manipulation
Supported
Pfand separation
Extracting deposit values distinct from the base product price
Supported
Nutritional table parsing
Structuring HTML tables into flat JSON keys
Supported
Weekly circular tracking
Extracting 'Aktionen' start and end dates
Supported
Change detection
Hash-based diffing for daily catalog sweeps
Supported
User order history
Requires authenticated customer sessions
Partial
DeutschlandCard points
Gated behind the loyalty program login wall
Partial
Infrastructure

Infrastructure powering the Netto pipeline

Open-source tooling on proven cloud infra — no vendor lock-in, full observability.

ScrapyPlaywrightPython 3.12RedisPostgreSQLApache AirflowAWS LambdaS3CloudWatch2CaptchaCapSolverResidential ProxiesDockerKubernetesGrafanaPrometheus
Scrapy + Playwright Stack

Scrapy handles crawl orchestration. Playwright handles JavaScript rendering, cookie consent, and zip code injection to access regional data.

DE Proxy Infrastructure

We maintain pools of residential ISP proxies across Germany. Rotation happens per request to ensure accurate regional pricing and avoid rate limits.

Cloud-Native Orchestration

Pipelines run on AWS Lambda and ECS. Airflow handles scheduling, dependency management, and SLA alerting for daily price sweeps.

Output & Delivery

Your data, your destination

Data delivered to where your team already works — no new tooling required.

JSON
Newline-delimited or nested structures
CSV
Flat file with typed columns for quick analysis
XLS
Excel compatible exports for commercial teams
Parquet
Columnar format for BigQuery and Snowflake
AWS S3
Direct bucket delivery on pipeline completion
Webhook
HTTP POST per record for real-time processing
API
REST endpoints for on-demand querying
BigQuery
Streamed directly into your dataset
S3
Direct bucket delivery — compatible with any data lake
// faq

Common questions.

About netto-online.de scraping, legality, and pipeline operations.

Ask us directly →
Is scraping Netto legal?

Scraping publicly available grocery data is generally permissible. DataFlirt extracts only public product, pricing, and nutritional data. We do not extract personal data or circumvent authentication walls.

How do you handle regional pricing?

We inject specific PLZ or zip codes into the session state using Playwright to extract localised prices and stock availability.

Can you extract nutritional info and allergens?

Yes, we parse the structured tables on product pages into flat JSON fields, capturing macronutrients, ingredients, and mandatory allergen warnings.

How fresh is the data?

Weekly specials are updated daily. Full catalog refreshes run on a 24-hour cycle to ensure accurate base pricing.

Do you separate Pfand from the product price?

Yes, deposit values are extracted as distinct fields alongside the base price to allow accurate total cost calculations.

What happens when Netto changes its layout?

Our selectors have multi-layer fallback chains. We monitor for schema drift and patch selectors before data drops occur.

$ dataflirt scope --new-project --source=netto-online.de ready

Tell us what
to extract.
We do the rest.

20-minute scoping call. Pilot dataset within the week. Production within two. Whether you need a daily price-monitoring feed or a one-off nutritional database dump, we scope, build, and operate the pipeline. Tell us what you need.

hello@dataflirt.com · Bengaluru · IST · typical reply < 4h
Related Scrapers

More in food drink and kitchen

Services

Data Extraction for Every Industry

View All Services →