SYSTEM all green source mecca.com.au queue 18,401 pages p99 latency 184ms dataflirt.com · scraper/mecca-com.au
RUN - 38 active pipelines - mecca.com.au live

Mecca beauty data,
at warehouse scale.

We extract product catalogues, shade variations, ingredient lists, stock levels, and customer reviews from Mecca. Delivered as clean JSON, CSV, or Parquet to S3, BigQuery, or Snowflake on your cadence.

Products extracted
18.4K /day
Stock updates
72.1K /24h
Review records
1.2M /run
Active pipelines
38
Uptime
99.98%
Data Dictionary

Every field we extract from mecca.com.au

Structured, schema-consistent data across all major object types — delivered clean, typed, and ready to query.

Complete list of extractable fields for Product Listings objects from mecca.com.au. All fields typed and schema-versioned.

skuproduct_namebrandcategorysub_categorypricecurrencyratingreview_countdescriptionhow_to_useimage_urls
product_listings
● 200 OK
"sku": "I-054321",
"product_name": "Protini Polypeptide Cream",
"brand": "Drunk Elephant",
"category": "Skincare",
"price": 112.0,
"currency": "AUD",
"rating": 4.6,
"review_count": 4812,
"in_stock": true
# skuproduct_namebrandcategorysub_categoryprice
1
2
3

Complete list of extractable fields for Pricing & Stock objects from mecca.com.au. All fields typed and schema-versioned.

skupricelist_pricecurrencyin_stockstock_statusonline_onlylimited_editionmax_purchase_qtystock_timestamp
pricing_& stock
● 200 OK
"sku": "I-054321",
"price": 112.0,
"list_price": 112.0,
"currency": "AUD",
"in_stock": true,
"online_only": false,
"limited_edition": false,
"stock_timestamp": "2026-05-12T09:14:00Z"
# skupricelist_pricecurrencyin_stockstock_status
1
2
3

Complete list of extractable fields for Ingredients & Specs objects from mecca.com.au. All fields typed and schema-versioned.

skuingredients_textclean_beauty_flagvegan_flagcruelty_freesize_volumeformatskin_type_suitabilityfinish
ingredients_& specs
● 200 OK
"sku": "I-054321",
"clean_beauty_flag": true,
"vegan_flag": true,
"cruelty_free": true,
"size_volume": "50ml",
"skin_type_suitability": "All Skin Types",
"finish": "Natural"
# skuingredients_textclean_beauty_flagvegan_flagcruelty_freesize_volume
1
2
3

Complete list of extractable fields for Shades & Variants objects from mecca.com.au. All fields typed and schema-versioned.

parent_skuvariant_skushade_nameshade_descriptioncolour_familyhex_codestock_statuspriceimage_url
shades_& variants
● 200 OK
"parent_sku": "I-041234",
"variant_sku": "V-041235",
"shade_name": "Mont Blanc",
"shade_description": "Light with neutral undertones",
"colour_family": "Fair",
"stock_status": "In Stock",
"price": 78.0
# parent_skuvariant_skushade_nameshade_descriptioncolour_familyhex_code
1
2
3

Complete list of extractable fields for Reviews & Ratings objects from mecca.com.au. All fields typed and schema-versioned.

review_idskureviewer_nicknamestar_ratingreview_titlereview_texthelpful_votesskin_typeage_rangerecommendedreview_date
reviews_& ratings
● 200 OK
"review_id": "REV-987654",
"sku": "I-054321",
"star_rating": 5,
"review_title": "Holy grail moisturiser",
"helpful_votes": 42,
"skin_type": "Combination",
"age_range": "25-34",
"recommended": true,
"review_date": "2026-04-18"
# review_idskureviewer_nicknamestar_ratingreview_titlereview_text
1
2
3

Capabilities

Everything you need from Mecca - nothing you don't

Our Mecca scraper handles every layer of the platform: brand catalogues, dynamic stock indicators, complex shade matrices, and the review corpus - with JavaScript rendering and anti-bot circumvention built in.

Full Product Catalogue Extraction

Title, description, how-to-use instructions, ingredients, size, and every metadata field Mecca surfaces - scraped at SKU level.

Complex Shade Mapping

Capture parent-child relationships for foundations and concealers, including shade names, descriptions, hex codes, and individual stock status.

Real-Time Stock Tracking

Monitor online availability, limited edition flags, and out-of-stock indicators - timestamped per crawl.

Review & Sentiment Mining

Full review text, star ratings, helpful vote counts, reviewer skin type, and age range - paginated across all review pages.

Ingredient & Formulation Data

Extract raw ingredient lists and parse clean beauty, vegan, and cruelty-free flags for product analysis.

Brand & Category Taxonomies

Map products to their exact category tree and brand portfolio to track brand dominance across the site.

Price Tracking

Capture current price in AUD or NZD, tracking any adjustments over time for competitor benchmarking.

Scheduled & Streaming Modes

Run one-off bulk exports or configure continuous pipelines at daily or weekly cadences with change-detection diffing.

Media Asset Extraction

Capture high-resolution product imagery, swatch photos, and video URLs associated with each SKU.

// engagement pipeline

From brand list to warehouse record

Brief in. Clean data out.

Define Scope
d 0

Provide brand lists, category URLs, or specific SKUs. We design the extraction schema together.

Pipeline Build
d 2–4

We configure Scrapy / Playwright crawlers, proxy rotation, session management, and CAPTCHA handling for mecca.com.au.

Validation & QA
d 4–6

Schema validation, null-rate checks, price-outlier detection, and sample reviews before full launch.

Delivery
ongoing

JSON / CSV / Parquet pushed to your S3 bucket, BigQuery dataset, or Snowflake stage on agreed cadence.

Under the hood

How our Mecca pipeline handles the hard parts

Premium retailers invest heavily in bot protection. Here is how we stay resilient - and why teams choose managed infrastructure over DIY.

pipeline-monitor · mecca.com.au · live ● active
// fingerprinting
Identity rotation
TLS fingerprintrandomised
User-agentrotated
IP poolresidential
Challenges blocked0
// pagination
Page coverage
48,291 pages queued running
// observability
Pipeline health
99.9%
uptime
142ms
p99 lat
0.3%
null rate
2
alerts
Anti-bot layer
Residential proxy rotation + fingerprint spoofing

Retail sites use advanced bot detection based on TLS fingerprints and IP reputation. Our crawlers use residential ISP proxies from AU/NZ pools with realistic browser fingerprints and full cookie session management.

JavaScript rendering
Full Playwright execution for SPA content

Mecca's product pages and shade selectors are heavily JavaScript-rendered. We run full Playwright browser sessions with JavaScript execution and lazy-load triggering to capture variant data accurately.

Variant complexity
Handling multi-dimensional shade matrices

Beauty products often have dozens of shades, each with unique stock states and descriptions. Our logic maps these parent-child relationships precisely, ensuring no variant is missed.

Review pagination
Deep crawling of third-party review widgets

Reviews are often loaded via third-party APIs. We intercept these network requests or paginate through the rendered DOM to extract the complete historical review corpus for every SKU.

Change detection
Only re-scrape what has changed

For large catalogues, we maintain a hash index of last-seen values per field. Subsequent runs only push diffs - reducing compute cost and downstream processing load.

Applications

Who uses Mecca data - and how

Teams across industries use mecca.com.au data to build competitive products and smarter operations.

01
Competitor Price Benchmarking

Beauty brands and competing retailers monitor Mecca's pricing strategies and brand assortment to adjust their own positioning.

02
Market Research & Trends

Analysts track new product launches, limited edition sell-out rates, and category expansion to identify beauty trends in the ANZ market.

03
Ingredient Analysis

Formulators and cosmetic chemists extract ingredient lists to track the adoption of specific actives and clean beauty standards.

04
Sentiment & Review Analysis

Brands mine customer reviews across their products to identify common complaints, packaging issues, or highly praised formulations.

05
Stock & Assortment Monitoring

Supply chain teams track out-of-stock rates for key brands to understand demand velocity and supply constraints.

06
Grey Market Detection

Global brands audit authorised retailer catalogues to ensure product ranges and pricing align with regional distribution agreements.

Why DataFlirt

"Mecca holds the most comprehensive dataset on premium beauty trends in the ANZ region - but extracting it requires navigating aggressive anti-bot protection and complex variant structures."

Most teams underestimate the investment required: reliable Mecca scraping requires residential proxies, full JavaScript rendering for shade selectors, CAPTCHA handling, daily selector maintenance, and anomaly monitoring. DataFlirt absorbs that complexity so your engineers can focus on the analysis - not the infrastructure.

Technical Spec

Mecca scraper - technical capabilities

Everything supported by our mecca.com.au scraper — rendered SPA elements, auth walls, rate-limit evasion and beyond.

JavaScript rendering
Full Playwright sessions - required for shade selectors and dynamic stock
Supported
CAPTCHA bypass
Automated 2Captcha + CapSolver integration for WAF challenges
Supported
Residential proxy rotation
ISP-grade residential IPs from AU/NZ pools - rotated per request
Supported
Shade mapping
Parent to child SKU relationships with all colour option combinations
Supported
Review pagination
Full review corpus including all pages and demographic metadata
Supported
Change detection (diffs)
Hash-based diff: only emit records with changed fields since last run
Supported
Webhook delivery
HTTP POST per record or batch - useful for downstream workflows
Supported
Beauty Loop rewards
Gated loyalty program data requires authenticated sessions
Partial
User purchase history
Private account data is strictly out of scope
Partial
Infrastructure

Infrastructure powering the Mecca pipeline

Open-source tooling on proven cloud infra — no vendor lock-in, full observability.

ScrapyPlaywrightPython 3.12RedisPostgreSQLApache AirflowAWS LambdaS3CloudWatch2CaptchaCapSolverResidential ProxiesDockerKubernetesGrafanaPrometheus
Scrapy + Playwright Stack

Scrapy handles crawl orchestration, deduplication, and retry logic. Playwright handles JavaScript rendering, cookie sessions, and interaction flows for complex shade selectors.

Residential Proxy Infrastructure

We maintain pools of residential ISP proxies across AU/NZ regions. Rotation happens per-request with sticky sessions where required to bypass strict WAF rules.

Cloud-Native Orchestration

Pipelines run on AWS Lambda and ECS. Airflow handles scheduling, dependency management, and SLA alerting. All state stored in managed Postgres.

Output & Delivery

Your data, your destination

Data delivered to where your team already works — no new tooling required.

JSON
Newline-delimited or nested - schema versioned per run
CSV
Flat file with typed columns - Excel/Sheets compatible
XLS
Standard Excel format for business analyst workflows
Parquet
Columnar format for BigQuery, Snowflake, Athena
AWS S3
Direct bucket delivery - compatible with any data lake
Webhook
HTTP POST per record for immediate downstream processing
API
REST endpoints to query your extracted datasets on demand
PostgreSQL
Upsert into your existing schema with conflict resolution
S3
Direct bucket delivery — compatible with any data lake
// faq

Common questions.

About mecca.com.au scraping, legality, and pipeline operations.

Ask us directly →
Is scraping Mecca legal?

Scraping publicly available information from Mecca is generally permissible. DataFlirt targets only public, non-authenticated product, pricing, and review data. We do not extract personal data, circumvent authentication walls, or scrape Beauty Loop member-only areas.

How do you handle Mecca's anti-bot systems?

We use residential ISP proxies from Australia and New Zealand, full Playwright browser sessions with realistic fingerprints, and request timing modelled on human behaviour. We monitor for block rate spikes in real time and trigger pool rotation automatically.

Can you extract all shades for a foundation?

Yes. Our pipeline maps the parent product to every individual shade variant, capturing the specific hex code, shade name, description, and stock status for each.

How fresh is the stock data?

Full catalogue refreshes at a daily cadence complete within a 4-6 hour window. For specific high-priority SKUs, we can configure higher frequency polling to detect out-of-stock events.

Can you extract ingredient lists?

Yes. We capture the full raw ingredient text block as presented on the product page, along with any clean beauty, vegan, or cruelty-free flags.

Do you support Mecca review scraping?

Yes. We extract the full review corpus including pagination across all reviews. Each record includes rating, title, body, helpful votes, and reviewer metadata like skin type and age range.

What is the minimum viable engagement?

Our smallest packages start at a defined brand list or category subset with weekly delivery. Contact us with your use case for a scoped quote.

$ dataflirt scope --new-project --source=mecca.com.au ready

Tell us what
to extract.
We do the rest.

20-minute scoping call. Pilot dataset within the week. Production within two. Whether you need a one-off catalogue dump or a continuous stock-monitoring feed across 18K products - we scope, build, and operate the pipeline. Tell us what you need.

hello@dataflirt.com · Bengaluru · IST · typical reply < 4h
Related Scrapers

More in beauty and skincare

Services

Data Extraction for Every Industry

View All Services →