We extract product catalogues, B2B pricing signals, stock availability, and technical specifications from Dustin.se. Delivered as clean JSON, CSV, or Parquet to S3, BigQuery, or Snowflake on your cadence.
Structured, schema-consistent data across all major object types — delivered clean, typed, and ready to query.
Complete list of extractable fields for Product Listings objects from dustin.se. All fields typed and schema-versioned.
"sku": "5011319245", "mpn": "21A00049MX", "ean": "0196801554231", "title": "Lenovo ThinkPad P16s Gen 1", "brand": "Lenovo", "price_ex_vat": 14995.0, "stock_status": "In stock", "delivery_time": "1-2 days"
| # | sku | mpn | ean | title | brand | category |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Pricing & Stock objects from dustin.se. All fields typed and schema-versioned.
"sku": "5011319245", "price_ex_vat": 14995.0, "list_price": 17495.0, "discount_pct": 14, "campaign_badge": "Weekly Deal", "stock_quantity": 42, "currency": "SEK", "scraped_at": "2026-05-12T09:14:00Z"
| # | sku | price_ex_vat | price_inc_vat | list_price | discount_pct | campaign_badge |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Technical Specs objects from dustin.se. All fields typed and schema-versioned.
"sku": "5011319245", "processor_family": "Intel Core i7", "ram_gb": 16, "storage_gb": 512, "storage_type": "SSD", "display_size": 16.0, "display_resolution": "1920 x 1200", "os": "Windows 11 Pro"
| # | sku | processor_family | ram_gb | storage_gb | storage_type | display_size |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Categories & Taxonomy objects from dustin.se. All fields typed and schema-versioned.
"category_id": "cat_10293", "l1_category": "Computers & Tablets", "l2_category": "Laptops", "l3_category": "Workstations", "product_count": 412, "breadcrumb": "Computers & Tablets > Laptops > Workstations"
| # | category_id | l1_category | l2_category | l3_category | url | product_count |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Accessories & Related objects from dustin.se. All fields typed and schema-versioned.
"main_sku": "5011319245", "accessory_sku": "5011284732", "relation_type": "Docking Station", "title": "Lenovo ThinkPad Universal USB-C Dock", "price_ex_vat": 1895.0, "stock_status": "In stock", "required_accessory": false
| # | main_sku | accessory_sku | relation_type | title | price_ex_vat | stock_status |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Our Dustin.se scraper handles complex B2B catalogues: dynamic VAT toggles, paginated technical specifications, real-time stock depth, and accessory mapping - all managed via resilient infrastructure.
Extract both Ex-VAT (B2B) and Inc-VAT (B2C) pricing models simultaneously, including volume discount tiers.
Capture processor models, RAM configurations, interface ports, and physical dimensions mapped to structured JSON keys.
Extract Manufacturer Part Numbers (MPN) and EAN codes for exact cross-referencing with other distributors.
Monitor stock availability flags, exact warehouse quantities, and estimated delivery lead times.
Extract recommended accessories, compatible docking stations, and service warranties linked to the primary SKU.
Identify active promotional badges, clearance items, and exact discount percentages across the catalogue.
Extract data across Dustin's regional domains (Sweden, Norway, Denmark, Finland) with local currency normalisation.
Receive only delta updates for pricing and stock changes, reducing processing overhead on your data warehouse.
Configure hourly stock checks for critical SKUs or daily full-catalogue refreshes based on your procurement needs.
Brief in. Clean data out.
Provide target categories, brand filters, or specific SKU lists. We design the extraction schema together.
We configure Playwright crawlers, regional proxies, session management, and specification parsers for dustin.se.
Schema validation, null-rate checks, price-outlier detection, and EAN format validation before full launch.
JSON / CSV / Parquet pushed to your S3 bucket, BigQuery dataset, or via Webhook on agreed cadence.
IT distributors employ strict rate limiting and complex DOM structures. Here is how we ensure reliable data delivery.
Dustin.se relies heavily on client-side rendering for pricing, stock updates, and accessory loading. We run full Playwright browser sessions to ensure dynamic widgets hydrate completely before extraction.
Pricing changes entirely based on the B2B/B2C toggle state. Our crawlers manage explicit cookie sessions to ensure we capture the correct VAT configuration and business-specific volume tiers without data contamination.
Aggressive scraping triggers IP bans and CAPTCHAs. We route requests through Swedish and Nordic residential proxies with human-like delay jitter to maintain continuous access to the catalogue.
IT hardware specs vary wildly between laptops and servers. We use dynamic key-value extraction logic to normalise technical specification tables into consistent JSON structures, regardless of product category.
Stock levels change by the minute. Our hashing engine compares current runs against historical state, outputting only SKUs where pricing or stock depth has shifted since the last crawl.
IT hardware resellers monitor Dustin's B2B pricing to adjust their own margins and maintain competitive positioning.
Enterprise procurement teams track historical pricing on standard hardware configurations to negotiate better enterprise contracts.
Hardware manufacturers (OEMs) audit Dustin.se to ensure compliance with Minimum Advertised Price (MAP) agreements.
Distributors track stock depth indicators across key categories to anticipate supply chain shortages and adjust inventory.
Retailers analyse Dustin's category taxonomy and new product introductions to identify gaps in their own IT hardware offerings.
Market analysts track campaign frequency, discount depth, and brand prominence to evaluate Nordic IT retail trends.
"Dustin.se holds the definitive catalogue for Nordic IT procurement, but extracting normalised B2B pricing and deep specs requires dedicated infrastructure."
Most teams underestimate the complexity of scraping IT hardware catalogues. Handling B2B versus B2C price toggles, paginated specification tables, and real-time stock indicators requires full browser rendering and session management. DataFlirt absorbs this complexity so your procurement and pricing teams can focus on analysis.
Everything supported by our dustin.se scraper — rendered SPA elements, auth walls, rate-limit evasion and beyond.
Open-source tooling on proven cloud infra — no vendor lock-in, full observability.
Scrapy handles crawl orchestration and deduplication. Playwright handles JavaScript rendering, VAT toggles, and interaction flows. Combined via scrapy-playwright middleware.
We maintain pools of residential ISP proxies across Sweden, Norway, Denmark, and Finland. Rotation happens per-request to avoid rate limits.
Pipelines run on AWS Lambda and ECS. Airflow handles scheduling, dependency management, and SLA alerting. All state stored in managed Postgres.
Data delivered to where your team already works — no new tooling required.
About dustin.se scraping, legality, and pipeline operations.
Ask us directly →Scraping publicly available pricing and product information from Dustin.se is generally permissible under applicable law for non-copyrighted factual data. DataFlirt targets only public, non-authenticated catalogue data. We do not extract personal data or circumvent authentication walls. Clients should consult legal counsel for specific commercial use cases.
We use regional residential ISP proxies, full Playwright browser sessions, and request timing modelled on human behaviour. We monitor for rate limiting in real time and trigger IP rotation automatically.
Yes. Our pipeline manages distinct session states to capture both Ex-VAT (B2B) and Inc-VAT (B2C) pricing simultaneously, alongside any volume-based discount tiers.
For targeted SKU lists, we can configure hourly stock checks. Full catalogue refreshes typically run on a daily cadence. We use change detection to deliver only the SKUs where stock status has changed.
No. DataFlirt extracts public catalogue pricing only. We do not handle authenticated sessions for customer-specific negotiated contract pricing.
Absolutely. We provide a sample run of up to 500 SKUs or specific categories as part of the pre-engagement scoping process so you can validate schema fit and data quality.
20-minute scoping call. Pilot dataset within the week. Production within two. Whether you need a one-off catalogue dump or a continuous price-monitoring feed across 100K SKUs - we scope, build, and operate the pipeline. Tell us what you need.