We extract product catalogues, shade variations, ingredient lists, stock levels, and customer reviews from Mecca. Delivered as clean JSON, CSV, or Parquet to S3, BigQuery, or Snowflake on your cadence.
Structured, schema-consistent data across all major object types — delivered clean, typed, and ready to query.
Complete list of extractable fields for Product Listings objects from mecca.com.au. All fields typed and schema-versioned.
"sku": "I-054321", "product_name": "Protini Polypeptide Cream", "brand": "Drunk Elephant", "category": "Skincare", "price": 112.0, "currency": "AUD", "rating": 4.6, "review_count": 4812, "in_stock": true
| # | sku | product_name | brand | category | sub_category | price |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Pricing & Stock objects from mecca.com.au. All fields typed and schema-versioned.
"sku": "I-054321", "price": 112.0, "list_price": 112.0, "currency": "AUD", "in_stock": true, "online_only": false, "limited_edition": false, "stock_timestamp": "2026-05-12T09:14:00Z"
| # | sku | price | list_price | currency | in_stock | stock_status |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Ingredients & Specs objects from mecca.com.au. All fields typed and schema-versioned.
"sku": "I-054321", "clean_beauty_flag": true, "vegan_flag": true, "cruelty_free": true, "size_volume": "50ml", "skin_type_suitability": "All Skin Types", "finish": "Natural"
| # | sku | ingredients_text | clean_beauty_flag | vegan_flag | cruelty_free | size_volume |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Shades & Variants objects from mecca.com.au. All fields typed and schema-versioned.
"parent_sku": "I-041234", "variant_sku": "V-041235", "shade_name": "Mont Blanc", "shade_description": "Light with neutral undertones", "colour_family": "Fair", "stock_status": "In Stock", "price": 78.0
| # | parent_sku | variant_sku | shade_name | shade_description | colour_family | hex_code |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Reviews & Ratings objects from mecca.com.au. All fields typed and schema-versioned.
"review_id": "REV-987654", "sku": "I-054321", "star_rating": 5, "review_title": "Holy grail moisturiser", "helpful_votes": 42, "skin_type": "Combination", "age_range": "25-34", "recommended": true, "review_date": "2026-04-18"
| # | review_id | sku | reviewer_nickname | star_rating | review_title | review_text |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Our Mecca scraper handles every layer of the platform: brand catalogues, dynamic stock indicators, complex shade matrices, and the review corpus - with JavaScript rendering and anti-bot circumvention built in.
Title, description, how-to-use instructions, ingredients, size, and every metadata field Mecca surfaces - scraped at SKU level.
Capture parent-child relationships for foundations and concealers, including shade names, descriptions, hex codes, and individual stock status.
Monitor online availability, limited edition flags, and out-of-stock indicators - timestamped per crawl.
Full review text, star ratings, helpful vote counts, reviewer skin type, and age range - paginated across all review pages.
Extract raw ingredient lists and parse clean beauty, vegan, and cruelty-free flags for product analysis.
Map products to their exact category tree and brand portfolio to track brand dominance across the site.
Capture current price in AUD or NZD, tracking any adjustments over time for competitor benchmarking.
Run one-off bulk exports or configure continuous pipelines at daily or weekly cadences with change-detection diffing.
Capture high-resolution product imagery, swatch photos, and video URLs associated with each SKU.
Brief in. Clean data out.
Provide brand lists, category URLs, or specific SKUs. We design the extraction schema together.
We configure Scrapy / Playwright crawlers, proxy rotation, session management, and CAPTCHA handling for mecca.com.au.
Schema validation, null-rate checks, price-outlier detection, and sample reviews before full launch.
JSON / CSV / Parquet pushed to your S3 bucket, BigQuery dataset, or Snowflake stage on agreed cadence.
Premium retailers invest heavily in bot protection. Here is how we stay resilient - and why teams choose managed infrastructure over DIY.
Retail sites use advanced bot detection based on TLS fingerprints and IP reputation. Our crawlers use residential ISP proxies from AU/NZ pools with realistic browser fingerprints and full cookie session management.
Mecca's product pages and shade selectors are heavily JavaScript-rendered. We run full Playwright browser sessions with JavaScript execution and lazy-load triggering to capture variant data accurately.
Beauty products often have dozens of shades, each with unique stock states and descriptions. Our logic maps these parent-child relationships precisely, ensuring no variant is missed.
Reviews are often loaded via third-party APIs. We intercept these network requests or paginate through the rendered DOM to extract the complete historical review corpus for every SKU.
For large catalogues, we maintain a hash index of last-seen values per field. Subsequent runs only push diffs - reducing compute cost and downstream processing load.
Beauty brands and competing retailers monitor Mecca's pricing strategies and brand assortment to adjust their own positioning.
Analysts track new product launches, limited edition sell-out rates, and category expansion to identify beauty trends in the ANZ market.
Formulators and cosmetic chemists extract ingredient lists to track the adoption of specific actives and clean beauty standards.
Brands mine customer reviews across their products to identify common complaints, packaging issues, or highly praised formulations.
Supply chain teams track out-of-stock rates for key brands to understand demand velocity and supply constraints.
Global brands audit authorised retailer catalogues to ensure product ranges and pricing align with regional distribution agreements.
"Mecca holds the most comprehensive dataset on premium beauty trends in the ANZ region - but extracting it requires navigating aggressive anti-bot protection and complex variant structures."
Most teams underestimate the investment required: reliable Mecca scraping requires residential proxies, full JavaScript rendering for shade selectors, CAPTCHA handling, daily selector maintenance, and anomaly monitoring. DataFlirt absorbs that complexity so your engineers can focus on the analysis - not the infrastructure.
Everything supported by our mecca.com.au scraper — rendered SPA elements, auth walls, rate-limit evasion and beyond.
Open-source tooling on proven cloud infra — no vendor lock-in, full observability.
Scrapy handles crawl orchestration, deduplication, and retry logic. Playwright handles JavaScript rendering, cookie sessions, and interaction flows for complex shade selectors.
We maintain pools of residential ISP proxies across AU/NZ regions. Rotation happens per-request with sticky sessions where required to bypass strict WAF rules.
Pipelines run on AWS Lambda and ECS. Airflow handles scheduling, dependency management, and SLA alerting. All state stored in managed Postgres.
Data delivered to where your team already works — no new tooling required.
About mecca.com.au scraping, legality, and pipeline operations.
Ask us directly →Scraping publicly available information from Mecca is generally permissible. DataFlirt targets only public, non-authenticated product, pricing, and review data. We do not extract personal data, circumvent authentication walls, or scrape Beauty Loop member-only areas.
We use residential ISP proxies from Australia and New Zealand, full Playwright browser sessions with realistic fingerprints, and request timing modelled on human behaviour. We monitor for block rate spikes in real time and trigger pool rotation automatically.
Yes. Our pipeline maps the parent product to every individual shade variant, capturing the specific hex code, shade name, description, and stock status for each.
Full catalogue refreshes at a daily cadence complete within a 4-6 hour window. For specific high-priority SKUs, we can configure higher frequency polling to detect out-of-stock events.
Yes. We capture the full raw ingredient text block as presented on the product page, along with any clean beauty, vegan, or cruelty-free flags.
Yes. We extract the full review corpus including pagination across all reviews. Each record includes rating, title, body, helpful votes, and reviewer metadata like skin type and age range.
Our smallest packages start at a defined brand list or category subset with weekly delivery. Contact us with your use case for a scoped quote.
20-minute scoping call. Pilot dataset within the week. Production within two. Whether you need a one-off catalogue dump or a continuous stock-monitoring feed across 18K products - we scope, build, and operate the pipeline. Tell us what you need.