We extract cosmetics listings, fragrance pricing signals, stock depth, brand intelligence, and customer reviews from Notino.it. Delivered as clean JSON, CSV, or Parquet to S3, BigQuery, or Snowflake on your cadence.
Structured, schema-consistent data across all major object types — delivered clean, typed, and ready to query.
Complete list of extractable fields for Product Listings objects from notino.it. All fields typed and schema-versioned.
"product_id": "NTN-8472", "title": "Armani My Way Eau de Parfum", "brand": "Armani", "price": 84.9, "currency": "EUR", "volume_ml": 50, "stock_status": "in_stock"
| # | product_id | title | brand | category | sub_category | price |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Pricing & Promos objects from notino.it. All fields typed and schema-versioned.
"product_id": "NTN-8472", "price": 84.9, "list_price": 115.0, "discount_pct": 26, "promo_code": "BEAUTY20", "promo_badge": "Black Friday Sale", "price_timestamp": "2023-11-24T08:12:00Z"
| # | product_id | price | list_price | discount_pct | promo_code | promo_badge |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Fragrance Notes objects from notino.it. All fields typed and schema-versioned.
"product_id": "NTN-8472", "fragrance_type": "Eau de Parfum", "top_notes": "['Bergamot', 'Orange Blossom']", "heart_notes": "['Tuberose', 'Jasmine']", "base_notes": "['Vanilla', 'White Musk', 'Cedarwood']", "gender_target": "Women"
| # | product_id | fragrance_type | top_notes | heart_notes | base_notes | fragrance_family |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Reviews & Ratings objects from notino.it. All fields typed and schema-versioned.
"review_id": "REV-99382", "product_id": "NTN-8472", "star_rating": 5, "review_title": "Profumo fantastico", "review_date": "2023-10-14", "skin_type": "Mista", "age_group": "25-34"
| # | review_id | product_id | reviewer_name | star_rating | review_title | review_body |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Brand Catalogues objects from notino.it. All fields typed and schema-versioned.
"brand_name": "Armani", "total_products": 342, "categories_covered": "['Profumi', 'Make-up', 'Viso']", "avg_price": 78.5, "avg_rating": 4.6, "brand_url": "https://www.notino.it/armani/"
| # | brand_id | brand_name | total_products | categories_covered | avg_price | avg_rating |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Our Notino scraper extracts the full cosmetics catalogue: pricing, stock availability, INCI ingredients, olfactory profiles, and customer reviews. We bypass regional blocks and bot protection natively.
Title, brand, volume variants, images, and every metadata field Notino surfaces across all beauty and fragrance categories.
Capture base price, discounted price, active promo codes, and gift-with-purchase flags. All timestamped per crawl.
Extract full ingredient lists for skincare and cosmetics to power formulation analysis and allergy-matching tools.
Structured extraction of top, heart, and base notes, plus fragrance families for perfumes and colognes.
Monitor stock status across all volume and colour variants to track inventory depth and stockouts.
Full review text, star ratings, and user attributes like skin type and age group. Paginated across all reviews.
Identify bundled products, advent calendars, and gift sets with component breakdown and implied discount calculation.
Scrape notino.it with Italian IP addresses to ensure accurate regional pricing, VAT inclusion, and local availability.
Run continuous pipelines at daily cadences with change-detection diffing to monitor fast-moving promo cycles.
Brief in. Clean data out.
Provide brand URLs, category paths, or keyword sets. We design the extraction schema together.
We configure Scrapy / Playwright crawlers, proxy rotation, and session management for notino.it.
Schema validation, null-rate checks, price-outlier detection, and sample reviews before full launch.
JSON / CSV / Parquet pushed to your S3 bucket, BigQuery dataset, or Snowflake stage on agreed cadence.
European eCommerce sites deploy strict rate limits and regional blocking. Here is how we maintain data flow without interruption.
Notino alters pricing and availability based on geographic IP. Our crawlers use Italian residential proxies to ensure scraped prices match what local consumers actually see.
Cosmetics often have dozens of shade variants, while perfumes have multiple volume options. We iterate through all variant selectors to capture specific pricing and stock status for each SKU.
Notino frequently uses banner-based promo codes. We parse DOM elements and promotional banners to map applicable codes directly to the product record.
Aggressive crawling triggers Cloudflare blocks. We manage request velocity, rotate user agents, and maintain healthy proxy pools to stay below rate-limit thresholds.
For large brand catalogues, we maintain a hash index of last-seen values per field. Subsequent runs only push diffs, reducing compute cost and downstream processing load.
Beauty retailers and pharmacies track Notino's aggressive pricing and promotional cycles to adjust their own pricing strategies.
Cosmetics brands monitor Notino to ensure MAP compliance and track unauthorised discounting.
Analysts track review velocity and category expansion to identify trending skincare active ingredients or fragrance notes.
Retail buyers analyse Notino's brand matrix and variant availability to identify gaps in their own product offerings.
Product development teams mine review text to understand customer complaints regarding formulation, packaging, or longevity.
Tracking out-of-stock rates across specific brands provides early signals regarding manufacturer supply chain bottlenecks.
"Notino holds the pulse of European beauty retail. Extracting its pricing and catalogue data provides an immediate competitive advantage in a low-margin sector."
Most retail teams underestimate the complexity of tracking cosmetics data at scale. Reliable Notino scraping requires handling complex variant matrices like shades and volumes, regional IP routing, and dynamic promotional logic. DataFlirt absorbs that complexity so your data engineers can focus on pricing algorithms and market analysis, not maintaining web scrapers.
Everything supported by our notino.it scraper — rendered SPA elements, auth walls, rate-limit evasion and beyond.
Open-source tooling on proven cloud infra — no vendor lock-in, full observability.
Scrapy handles crawl orchestration and retry logic. Playwright handles JavaScript rendering for dynamic variant loading and promo banners.
We maintain pools of Italian residential ISP proxies. Rotation happens per-request to ensure accurate local pricing and avoid geo-blocks.
Pipelines run on AWS Lambda and ECS. Airflow handles scheduling, dependency management, and SLA alerting. All state stored in managed Postgres.
Data delivered to where your team already works — no new tooling required.
About notino.it scraping, legality, and pipeline operations.
Ask us directly →Scraping publicly available information from Notino is generally permissible under applicable EU law. DataFlirt targets only public, non-authenticated product, pricing, and review data. We do not extract personal data or circumvent authentication walls.
We route all requests through Italian residential proxies. This ensures the site returns correct Italian pricing, VAT, and local stock availability rather than defaulting to a generic European view.
Yes. Our crawlers iterate through every variant selector on the product page, capturing the specific SKU, price, stock status, and image URL for each individual shade or volume.
Yes. We extract base prices, discounted prices, and parse promotional banners to identify active discount codes applicable to the product.
Full catalogue refreshes at daily cadence complete within a 4-8 hour window. For specific high-priority brands or categories, we can configure sub-hourly tracking.
Yes. We parse the product description and metadata to extract structured olfactory profiles, including top notes, heart notes, base notes, and fragrance family.
Absolutely. We provide a sample run of up to 500 products as part of the pre-engagement scoping process so you can validate schema fit and data quality.
20-minute scoping call. Pilot dataset within the week. Production within two. Whether you need a daily pricing feed across the entire fragrance catalogue or targeted competitor tracking, we scope, build, and operate the pipeline. Tell us what you need.