We extract yarn specifications, needle gauges, colourways, and pattern kits from Wollplatz.de. Delivered as clean JSON, CSV, or Parquet to S3, BigQuery, or Snowflake on your cadence.
Structured, schema-consistent data across all major object types — delivered clean, typed, and ready to query.
Complete list of extractable fields for Yarn Specifications objects from wollplatz.de. All fields typed and schema-versioned.
"product_id": "WP-10492", "brand": "Drops", "name": "Drops Alaska", "composition": "100% Schurwolle", "needle_size": "5 mm", "tension_gauge": "17 Maschen x 22 Reihen", "weight_grams": 50, "length_meters": 70, "price": 1.95, "currency": "EUR"
| # | product_id | brand | name | composition | needle_size | tension_gauge |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Colourways objects from wollplatz.de. All fields typed and schema-versioned.
"product_id": "WP-10492", "colour_code": "02", "colour_name": "Natur", "stock_status": "in_stock", "price_modifier": 0.0, "image_url": "https://www.wollplatz.de/images/drops-alaska-02.jpg", "scraped_at": "2023-10-24T08:15:00Z"
| # | product_id | colour_code | colour_name | stock_status | price_modifier | image_url |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Pattern Kits objects from wollplatz.de. All fields typed and schema-versioned.
"kit_id": "PK-8831", "name": "Babydecke Häkelpaket", "designer": "Wollplatz", "skill_level": "Anfänger", "included_yarns": "['Yarn and Colors Epic']", "required_tools": "['Häkelnadel 5mm']", "base_price": 24.5, "description": "Ein komplettes Paket für eine weiche Babydecke."
| # | kit_id | name | designer | skill_level | included_yarns | required_tools |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Pricing & Stock objects from wollplatz.de. All fields typed and schema-versioned.
"product_id": "WP-10492", "sku": "DR-AL-02", "base_price": 2.15, "discount_price": 1.95, "discount_pct": 9.3, "currency": "EUR", "in_stock": true, "delivery_time_days": "2-3", "scraped_at": "2023-10-24T08:15:00Z"
| # | product_id | sku | base_price | discount_price | discount_pct | currency |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Categories objects from wollplatz.de. All fields typed and schema-versioned.
"category_id": "CAT-WOLLE", "category_name": "Wolle & Garne", "parent_category": "Startseite", "product_count": 4821, "url": "https://www.wollplatz.de/wolle", "breadcrumb_path": "Startseite > Wolle & Garne", "scraped_at": "2023-10-24T08:15:00Z"
| # | category_id | category_name | parent_category | product_count | url | breadcrumb_path |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Our Wollplatz scraper handles the complex parent-child relationships of yarn lines, extracting every colourway, tension metric, and stock indicator using automated JavaScript rendering.
Extract composition percentages, needle sizes, tension gauges, and weight categories directly from product descriptions.
Capture every variant within a yarn line, including specific colour codes, names, and variant-specific stock levels.
Monitor availability statuses and delivery estimates across thousands of SKUs to forecast supply chain movements.
Track base prices, promotional discounts, and bulk-buy incentives across all brands and categories.
Break down pattern kits into their constituent parts: required yarns, tool specifications, and skill level indicators.
Index complete collections from Drops, Lana Grossa, Phildar, and Rico Design, tracking new additions.
Extract raw German product copy and metadata, preserving specific textile terminology for accurate downstream translation or analysis.
Receive only updated records for stock changes and price fluctuations, minimising processing overhead.
Extract dimensions, materials, and compatibility metrics for knitting needles, crochet hooks, and accessories.
Brief in. Clean data out.
Provide target categories, specific brands, or update frequencies. We design the extraction schema together.
We configure Scrapy crawlers, proxy rotation through EU IP pools, and variant mapping logic for wollplatz.de.
Schema validation, null-rate checks, and colourway count verification before full launch.
JSON / CSV / Parquet pushed to your S3 bucket, BigQuery dataset, or Snowflake stage on agreed cadence.
Extracting eCommerce textile data requires managing complex variant structures and dynamic frontend frameworks.
A single yarn product on Wollplatz often has over 50 colour variants. We map these parent-child relationships accurately, ensuring each colourway has its own distinct SKU record with corresponding stock and pricing.
Colour switchers and dynamic stock indicators on Wollplatz rely on JavaScript. We run Playwright sessions to hydrate the DOM and capture accurate variant states that static HTML scrapers miss.
Product specifications like tension gauge and needle size are often embedded in unstructured text or varied table formats. Our pipeline uses regex and NLP fallbacks to normalise these metrics into strict numerical fields.
To prevent geoblocking and ensure accurate localised pricing and stock data, we route requests through residential IP pools located in Germany and surrounding EU regions.
We maintain a hash index of last-seen values per field. Subsequent runs only push diffs — reducing compute cost and downstream processing load when tracking daily stock fluctuations.
Textile retailers track Wollplatz pricing and discount strategies across major brands like Drops and Lana Grossa to optimise their own margins.
Supply chain analysts monitor stock-out rates on specific colourways and yarn weights to predict seasonal crafting trends.
Wholesalers and independent retailers extract detailed composition and tension metrics to enrich their own product databases.
Brands analyse the proliferation of new yarn compositions (e.g., recycled fibres) and pattern kit themes to guide product development.
Crafting apps and platforms aggregate required materials and tool specifications from Wollplatz pattern kits to build comprehensive project databases.
Marketing teams track the velocity of new colourway releases and category expansions to identify emerging trends in the European crafting market.
"Wollplatz holds Europe's most structured dataset for yarn composition, tension metrics, and needle specifications — if you can parse the variants."
Extracting textile data requires mapping complex parent-child relationships where a single yarn line might have 80 colourways, varying stock levels, and fluctuating prices. DataFlirt handles the JavaScript rendering and DOM extraction so your engineers can focus on inventory modelling, not proxy rotation.
Everything supported by our wollplatz.de scraper — rendered SPA elements, auth walls, rate-limit evasion and beyond.
Open-source tooling on proven cloud infra — no vendor lock-in, full observability.
Scrapy handles crawl orchestration, deduplication, and retry logic. Playwright handles JavaScript rendering, cookie sessions, and interaction flows. Combined via scrapy-playwright middleware.
We maintain pools of residential ISP proxies across EU regions. Rotation happens per-request with sticky sessions where required. IP score monitoring prevents blacklisted pool contamination.
Pipelines run on AWS Lambda (burst) and ECS (sustained). Airflow handles scheduling, dependency management, and SLA alerting. All state stored in managed Postgres.
Data delivered to where your team already works — no new tooling required.
About wollplatz.de scraping, legality, and pipeline operations.
Ask us directly →Yes. Our scraper iterates through all colour variants on a product page, capturing the specific colour code, name, image, and stock status for each child SKU.
We extract the raw German text directly from the DOM. For critical metrics like needle size (Nadelstärke) and tension (Maschenprobe), our pipeline uses regex to parse out the numerical values into structured fields.
Yes. DataFlirt only extracts publicly available commercial data: product specifications, pricing, and stock levels. We do not scrape personally identifiable information (PII) or user accounts.
Yes. We can configure pipelines to run at high frequencies (e.g., daily or hourly) to monitor changes in stock status and delivery estimates across specific product categories.
No. We extract the metadata associated with pattern kits (required yarns, tools, skill level, description). We do not download or distribute copyrighted pattern PDFs.
We use resilient selectors with multiple fallback chains. If Wollplatz updates their DOM, our monitoring catches null-rate spikes immediately, and our engineers update the selectors within our SLA window.
20-minute scoping call. Pilot dataset within the week. Production within two. Whether you need a one-off catalogue extraction or continuous stock monitoring across 15,000 SKUs — we scope, build, and operate the pipeline. Tell us what you need.