SYSTEM all green source lornajane.com queue 4,192 pages p99 latency 184ms dataflirt.com · scraper/lornajane-com
RUN - 14 active pipelines - lornajane.com live

Lorna Jane data,
at warehouse scale.

We extract product listings, sizing matrices, markdown signals, fabric intelligence, and fit reviews from Lorna Jane. Delivered as clean JSON, CSV, or Parquet to S3, BigQuery, or Snowflake on your cadence.

Products extracted
1,850 /day
Price updates
4,200 /24h
Review records
32K /run
Active pipelines
14
Uptime
99.95%
Data Dictionary

Every field we extract from lornajane.com

Structured, schema-consistent data across all major object types — delivered clean, typed, and ready to query.

Complete list of extractable fields for Product Listings objects from lornajane.com. All fields typed and schema-versioned.

skutitlecategorysub_categorypricecurrencycolour_namefabric_typesupport_levelratingreview_countdescriptioncare_instructionsimage_urlsurl
product_listings
● 200 OK
"sku": "LB0321_BLK",
"title": "Compress & Compact Sports Bra",
"category": "Sports Bras",
"price": 75.0,
"currency": "AUD",
"colour_name": "Black",
"support_level": "High Support",
"rating": 4.8,
"review_count": 142
# skutitlecategorysub_categorypricecurrency
1
2
3

Complete list of extractable fields for Pricing & Promos objects from lornajane.com. All fields typed and schema-versioned.

skucurrent_priceoriginal_pricediscount_pctdiscount_abspromo_badgesale_categorycurrencyprice_timestamp
pricing_& promos
● 200 OK
"sku": "LB0321_BLK",
"current_price": 50.0,
"original_price": 75.0,
"discount_pct": 33.3,
"promo_badge": "Sale",
"sale_category": "End of Season",
"currency": "AUD",
"price_timestamp": "2026-05-12T09:14:00Z"
# skucurrent_priceoriginal_pricediscount_pctdiscount_abspromo_badge
1
2
3

Complete list of extractable fields for Inventory & Sizing objects from lornajane.com. All fields typed and schema-versioned.

skuparent_idsizecolourin_stocklow_stock_warningstock_levelrestock_datescraped_at
inventory_& sizing
● 200 OK
"sku": "LB0321_BLK_S",
"size": "Small",
"colour": "Black",
"in_stock": true,
"low_stock_warning": true,
"stock_level": "Low",
"scraped_at": "2026-05-12T09:14:00Z"
# skuparent_idsizecolourin_stocklow_stock_warning
1
2
3

Complete list of extractable fields for Reviews & Fit objects from lornajane.com. All fields typed and schema-versioned.

review_idskureviewer_namestar_ratingfit_ratingcomfort_ratingquality_ratingreview_titlereview_bodyreview_dateverified_buyer
reviews_& fit
● 200 OK
"review_id": "REV-98234",
"sku": "LB0321_BLK",
"star_rating": 5,
"fit_rating": "True to Size",
"review_title": "Best running bra",
"review_date": "2026-04-18",
"verified_buyer": true
# review_idskureviewer_namestar_ratingfit_ratingcomfort_rating
1
2
3

Complete list of extractable fields for Categories & Navigation objects from lornajane.com. All fields typed and schema-versioned.

category_idcategory_nameparent_categoryurlproduct_countbreadcrumb_pathfeatured_promotionscraped_at
categories_& navigation
● 200 OK
"category_name": "High Support Sports Bras",
"parent_category": "Sports Bras",
"url": "/collections/high-support-sports-bras",
"product_count": 45,
"breadcrumb_path": "Home > Sports Bras > High Support",
"scraped_at": "2026-05-12T09:14:00Z"
# category_idcategory_nameparent_categoryurlproduct_countbreadcrumb_path
1
2
3

Capabilities

Everything you need from Lorna Jane - nothing you don't

Our Lorna Jane scraper handles the complete retail matrix: product listings, complex size-and-colour variants, dynamic pricing, and fit reviews - with full JavaScript rendering built in.

Full Catalogue Extraction

Title, fabric composition, support levels, care instructions, and high-resolution image URLs - scraped across all activewear categories.

Size & Variant Mapping

Capture the complete matrix of available sizes and colours per parent SKU, including low stock indicators and out-of-stock flags.

Markdown & Price Tracking

Extract current price, original RRP, and discount percentages. Track promotional badges and seasonal sale inclusions.

Fit & Comfort Review Mining

Full review text, star ratings, and specific fit metrics (e.g. True to Size) to analyse garment performance and customer sentiment.

Regional Pricing Grids

Scrape location-specific catalogues to compare pricing and assortment across Australian, US, and UK storefronts.

Scheduled Diffs

Run continuous pipelines at daily or weekly cadences. We maintain state and only push records that have changed.

Asset Extraction

Collect CDN links for all product imagery, including front, back, detail, and model lifestyle shots.

Category & Assortment Data

Map the entire category tree and breadcrumb structure to understand merchandising hierarchy and product density.

Cross-sell Recommendations

Extract 'Wear it with' and 'You may also like' product associations directly from the product detail pages.

// engagement pipeline

From target category to warehouse record

Brief in. Clean data out.

Define Scope
d 0

Provide category URLs, specific SKUs, or geographic regions. We design the extraction schema together.

Pipeline Build
d 2–4

We configure Scrapy / Playwright crawlers, handle regional routing, and map the complex variant DOM structures.

Validation & QA
d 4–6

Schema validation, null-rate checks, and variant completeness testing before full launch.

Delivery
ongoing

JSON / CSV / Parquet pushed to your S3 bucket, BigQuery dataset, or Snowflake stage on agreed cadence.

Under the hood

How our retail pipeline handles the hard parts

Modern eCommerce sites rely heavily on dynamic hydration and complex variant matrices. Here is how we extract clean data without breaking.

pipeline-monitor · lornajane.com · live ● active
// fingerprinting
Identity rotation
TLS fingerprintrandomised
User-agentrotated
IP poolresidential
Challenges blocked0
// pagination
Page coverage
48,291 pages queued running
// observability
Pipeline health
99.9%
uptime
142ms
p99 lat
0.3%
null rate
2
alerts
JavaScript rendering
Full Playwright execution for SPA content

Product pages and variant selectors are heavily JavaScript-rendered. We run full Playwright browser sessions to trigger size and colour selections, capturing dynamic price and stock changes that headless HTTP clients miss entirely.

Variant normalisation
Flattening the size and colour matrix

Retailers nest SKUs deeply. Our extractors flatten parent-child relationships into clean, queryable rows, ensuring every size and colour combination is represented with its specific stock status and price.

Anti-bot layer
Residential proxy rotation

High-frequency scraping triggers rate limits. Our crawlers use residential ISP proxies with realistic browser fingerprints and request timing to blend in with legitimate shopper traffic.

Schema stability
Resilient selectors with fallback chains

eCommerce layouts change during major sales events. Our selector strategy uses fallback chains - CSS selectors, XPath, and JSON-LD extraction - so a layout change does not break your data pipeline overnight.

Change detection
Only re-scrape what has changed

We maintain a hash index of last-seen values per SKU. Subsequent runs only push diffs - reducing compute cost and downstream processing load. You get a clean changelog of markdowns and stockouts.

Applications

Who uses Lorna Jane data - and how

Teams across industries use lornajane.com data to build competitive products and smarter operations.

01
Pricing & Markdown Intel

Retailers monitor competitor promotional cadences, discount depths, and end-of-season sale timing to optimise their own pricing strategies.

02
Assortment Planning

Merchandising teams analyse category density, colour trends, and new product introductions to inform their buying and design cycles.

03
Trend Analysis

Apparel analysts track technical fabric adoption, support level distributions, and silhouette changes in the activewear market.

04
Size & Fit Optimisation

Product teams mine customer reviews for fit complaints and sizing inconsistencies to improve their own garment grading.

05
Stockout Monitoring

Supply chain analysts track inventory depletion rates across specific sizes to estimate sales velocity and demand curves.

06
Brand Benchmarking

Investors and analysts track catalogue size, review velocity, and markdown frequency to evaluate brand health and market positioning.

Why DataFlirt

"Lorna Jane's activewear catalogue contains critical signals on technical fabric trends, sizing distributions, and markdown cadences - data that requires persistent extraction."

Fashion scrapers fail on variant explosion and dynamic inventory states. We handle the complex matrix of sizes, colours, and regional pricing grids. DataFlirt manages the proxy rotation and DOM parsing so your team can focus on merchandising analytics rather than pipeline maintenance.

Technical Spec

Lorna Jane scraper - technical capabilities

Everything supported by our lornajane.com scraper — rendered SPA elements, auth walls, rate-limit evasion and beyond.

JavaScript rendering
Full Playwright sessions required for dynamic variant selection and stock checks
Supported
CAPTCHA bypass
Automated 2Captcha + CapSolver integration for rate-limit blocks
Supported
Residential proxy rotation
ISP-grade residential IPs from AU / US / UK pools
Supported
Variant matrix mapping
Parent to child SKU relationships mapping every size and colour combination
Supported
Review pagination
Extraction of all paginated reviews including fit and comfort metrics
Supported
Regional pricing
Extraction across specific geographic storefronts (e.g. AU vs US)
Supported
Change detection (diffs)
Hash-based diff: only emit records with changed fields since last run
Supported
Webhook delivery
HTTP POST per record or batch for real-time stockout alerts
Supported
User Wishlists
Gated data tied to individual authenticated user accounts
Partial
Checkout & Cart Data
Live shipping rates and cart-specific promotional applications
Partial
Infrastructure

Infrastructure powering the retail pipeline

Open-source tooling on proven cloud infra — no vendor lock-in, full observability.

ScrapyPlaywrightPython 3.12RedisPostgreSQLApache AirflowAWS LambdaS3CloudWatch2CaptchaCapSolverResidential ProxiesDockerKubernetesGrafanaPrometheus
Scrapy + Playwright Stack

Scrapy handles crawl orchestration and deduplication. Playwright handles JavaScript rendering, cookie sessions, and the complex variant selection flows required for modern apparel sites.

Residential Proxy Infrastructure

We maintain pools of residential ISP proxies across target regions. Rotation happens per-request with sticky sessions to ensure consistent regional pricing and currency data.

Cloud-Native Orchestration

Pipelines run on AWS Lambda and ECS. Airflow handles scheduling, dependency management, and SLA alerting. All state stored in managed Postgres.

Output & Delivery

Your data, your destination

Data delivered to where your team already works — no new tooling required.

JSON
Newline-delimited or nested - schema versioned per run
CSV
Flat file with typed columns - Excel/Sheets compatible
XLS
Excel format for direct analyst consumption
Parquet
Columnar format for BigQuery, Snowflake, Athena
AWS S3
Direct bucket delivery - compatible with any data lake
Webhook
HTTP POST per record for real-time downstream processing
API
REST endpoints to query historical pricing and stock data
BigQuery
Streamed directly into your dataset with schema auto-detect
S3
Direct bucket delivery — compatible with any data lake
// faq

Common questions.

About lornajane.com scraping, legality, and pipeline operations.

Ask us directly →
Is scraping Lorna Jane legal?

Scraping publicly available information from retail websites is generally permissible under applicable law. DataFlirt targets only public product, pricing, and review data. We do not extract personal data or circumvent authentication walls. Clients should consult legal counsel for specific use cases.

How do you handle the complex size and colour variants?

Our pipelines use Playwright to interact with the DOM, systematically selecting every available size and colour combination to extract the specific price, SKU, and stock status for each variant. The data is then normalised into a flat, queryable structure.

Can you track pricing across different regions?

Yes. We route requests through region-specific residential proxies to capture accurate localised pricing, currency, and availability for Australian, US, UK, and other international storefronts.

How fresh is the inventory data?

We configure pipeline frequency based on your requirements. Daily sweeps capture broad markdown trends, while high-frequency intra-day runs can monitor stockouts on high-velocity SKUs.

Do you extract product reviews and fit metrics?

Yes. We paginate through all product reviews, capturing star ratings, review text, and specific fit metrics such as 'True to Size' or 'Runs Small' to provide a complete view of customer sentiment.

Can I request a sample dataset before committing?

Absolutely. We provide a sample run of up to 100 SKUs as part of the pre-engagement scoping process so you can validate schema fit, variant completeness, and data quality before signing any contract.

$ dataflirt scope --new-project --source=lornajane.com ready

Tell us what
to extract.
We do the rest.

20-minute scoping call. Pilot dataset within the week. Production within two. Whether you need a one-off catalogue dump or a continuous price-monitoring feed across thousands of SKUs - we scope, build, and operate the pipeline. Tell us what you need.

hello@dataflirt.com · Bengaluru · IST · typical reply < 4h
Related Scrapers

More in fashion and apparel

Services

Data Extraction for Every Industry

View All Services →