SYSTEM all green source zulily.com queue 12,481 events p99 latency 184ms dataflirt.com · scraper/zulily-com
RUN · 42 active pipelines · zulily.com live

Zulily deal data,
extracted at scale.

We extract flash sale events, boutique pricing, inventory levels, and apparel specifications from Zulily. Delivered as clean JSON, CSV, or Parquet to S3 or Snowflake on your cadence.

Products extracted
184K /day
Deal updates
42K /12h
Brand events
1,294 /run
Active pipelines
42
Uptime
99.95%
Data Dictionary

Every field we extract from zulily.com

Structured, schema-consistent data across all major object types — delivered clean, typed, and ready to query.

Complete list of extractable fields for Flash Sale Events objects from zulily.com. All fields typed and schema-versioned.

event_idbrand_nameevent_titlestart_timeend_timebanner_image_urlproduct_countcategorystatusscraped_at
flash_sale events
● 200 OK
"event_id": "evt_948172",
"brand_name": "Matilda Jane",
"event_title": "Matilda Jane Clothing: Up to 60% Off",
"start_time": "2023-10-14T06:00:00Z",
"end_time": "2023-10-17T06:00:00Z",
"product_count": 142,
"status": "active"
# event_idbrand_nameevent_titlestart_timeend_timebanner_image_url
1
2
3

Complete list of extractable fields for Product Listings objects from zulily.com. All fields typed and schema-versioned.

product_idevent_idtitlebrandcategorysub_categorypricemsrpdiscount_pctcoloursize_optionsmaterialcare_instructionsimage_urls
product_listings
● 200 OK
"product_id": "prd_883192",
"title": "Navy Floral Tunic",
"brand": "Matilda Jane",
"price": 24.99,
"msrp": 48.0,
"discount_pct": 47,
"colour": "Navy",
"material": "95% Cotton, 5% Spandex"
# product_idevent_idtitlebrandcategorysub_category
1
2
3

Complete list of extractable fields for Pricing & Inventory objects from zulily.com. All fields typed and schema-versioned.

product_idcurrent_priceoriginal_pricecurrencyin_stockstock_levelwaitlist_availableshipping_estimatereturn_policyprice_timestamp
pricing_& inventory
● 200 OK
"product_id": "prd_883192",
"current_price": 24.99,
"original_price": 48.0,
"currency": "USD",
"in_stock": true,
"waitlist_available": false,
"shipping_estimate": "10-14 days",
"price_timestamp": "2023-10-15T08:30:00Z"
# product_idcurrent_priceoriginal_pricecurrencyin_stockstock_level
1
2
3

Complete list of extractable fields for Apparel Specifications objects from zulily.com. All fields typed and schema-versioned.

product_idsize_chart_urlfit_typemodel_measurementsfabric_compositionpatternnecklinesleeve_lengthcare_instructions
apparel_specifications
● 200 OK
"product_id": "prd_883192",
"fit_type": "Relaxed",
"fabric_composition": "95% Cotton, 5% Spandex",
"pattern": "Floral",
"neckline": "Crew",
"sleeve_length": "Long Sleeve",
"care_instructions": "Machine wash cold"
# product_idsize_chart_urlfit_typemodel_measurementsfabric_compositionpattern
1
2
3

Complete list of extractable fields for Brand Data objects from zulily.com. All fields typed and schema-versioned.

brand_idbrand_nameactive_eventspast_eventsaverage_discountbrand_descriptionlogo_urlcategory_focusscraped_at
brand_data
● 200 OK
"brand_id": "brd_441",
"brand_name": "Matilda Jane",
"active_events": 2,
"average_discount": 45.5,
"category_focus": "Girls Apparel",
"scraped_at": "2023-10-15T08:30:00Z"
# brand_idbrand_nameactive_eventspast_eventsaverage_discountbrand_description
1
2
3

Capabilities

Everything you need from Zulily, nothing you don't

Our Zulily scraper handles every layer of the platform: flash sale events, boutique pricing, variant matrices, and waitlist status, with JavaScript rendering and session management built in.

Flash Sale Monitoring

Track event start and end times, banner images, and total product counts per boutique event.

Dynamic Pricing Extraction

Capture boutique pricing, MSRP, and calculated discount percentages timestamped per crawl.

Variant Matrix Mapping

Map colour and size combinations accurately across apparel listings to build complete product matrices.

Inventory & Waitlist Tracking

Monitor stock depletion and waitlist availability indicators for high-demand boutique items.

Brand Event Aggregation

Group products by daily deal events and extract brand-level metadata and historical event frequency.

Image & Media Scraping

Extract high-resolution product image URLs and brand banner assets for visual analysis.

Size Chart Extraction

Structured extraction of sizing tables, fit descriptions, and fabric composition data.

Category Navigation

Crawl through women's, kids, home, and beauty categories systematically.

Anti-Bot Circumvention

Bypass perimeter defenses and aggressive login prompts using residential proxies and realistic fingerprints.

Scheduled Delivery

Hourly or daily syncs matching Zulily's morning deal launch cadence for maximum data freshness.

// engagement pipeline

From event list to warehouse record

Brief in. Clean data out.

Define Scope
d 0

Provide target categories, brand names, or event URLs. We design the extraction schema together.

Pipeline Build
d 2–4

We configure Scrapy and Playwright crawlers, proxy rotation, and session management for zulily.com.

Validation & QA
d 4–6

Schema validation, null-rate checks, price-outlier detection, and sample data before full launch.

Delivery
ongoing

JSON, CSV, or Parquet pushed to your S3 bucket, BigQuery dataset, or Snowflake stage on agreed cadence.

Under the hood

How our Zulily pipeline handles the hard parts

Zulily's ephemeral catalogue and aggressive login walls require specialised infrastructure. Here is how we maintain stable extraction.

pipeline-monitor · zulily.com · live ● active
// fingerprinting
Identity rotation
TLS fingerprintrandomised
User-agentrotated
IP poolresidential
Challenges blocked0
// pagination
Page coverage
48,291 pages queued running
// observability
Pipeline health
99.9%
uptime
142ms
p99 lat
0.3%
null rate
2
alerts
Flash Sale Cadence
Burst crawling at deal launch

Zulily launches thousands of products simultaneously at 6 AM PST. We scale our Kubernetes workers dynamically to capture the entire catalogue within minutes of launch before high-demand items sell out.

Auth-Wall Bypassing
Session management and cookie rotation

Zulily frequently gates browsing behind a login wall. We manage authenticated cookie pools and rotate sessions automatically to maintain uninterrupted access to public catalogue data.

Variant Complexity
Multi-dimensional size and colour grids

Apparel listings feature complex matrices of sizes and colours. Our parsers map every possible combination, recording specific stock status and pricing for each SKUs variant.

Image Rendering
Triggering lazy-loaded galleries

Product images are heavily lazy-loaded. We use Playwright to simulate scroll behaviour and network idle states, ensuring all high-resolution image URLs are captured before extraction.

Anti-Bot Evasion
Residential proxies for IP reputation

We route requests through US-based residential ISP proxies to avoid datacenter IP bans, maintaining high success rates even during aggressive burst crawls.

Applications

Who uses Zulily data and how

Teams across industries use zulily.com data to build competitive products and smarter operations.

01
Competitor Price Benchmarking

Retailers monitor Zulily's boutique pricing and discount depths to adjust their own promotional strategies.

02
Brand MAP Monitoring

Brands track flash sales to ensure third-party sellers and liquidators adhere to Minimum Advertised Price agreements.

03
Inventory Trend Analysis

Analysts monitor waitlist metrics and stock depletion rates to gauge consumer demand for specific apparel categories.

04
Flash Sale Strategy

E-commerce strategists study event duration, product mix, and timing to optimise their own daily deal structures.

05
Market Research

Firms track emerging boutique brands and category saturation to identify new wholesale opportunities.

06
AI Fashion Models

Machine learning teams use structured apparel attributes and high-resolution images to train visual search and recommendation engines.

Why DataFlirt

"Zulily's flash sale model creates a highly volatile catalogue where thousands of products vanish daily. Capturing this requires precision timing."

Extracting data from ephemeral daily deals requires infrastructure that can burst at specific hours. We handle the residential proxies, JavaScript rendering, and session management needed to bypass login walls and extract complete apparel catalogues before the event timer hits zero.

Technical Spec

Zulily scraper: technical capabilities

Everything supported by our zulily.com scraper — rendered SPA elements, auth walls, rate-limit evasion and beyond.

JavaScript rendering
Full Playwright sessions required for lazy-loaded images and dynamic variant selection
Supported
CAPTCHA bypass
Automated 2Captcha and CapSolver integration
Supported
Residential proxy rotation
ISP-grade residential IPs from US pools rotated per request
Supported
Flash sale timer tracking
Exact start and end timestamps extracted for every event
Supported
Size/Colour variant mapping
Full matrix extraction of all available apparel options
Supported
High-res image URL extraction
Capture of uncompressed product gallery images
Supported
Change detection (diffs)
Hash-based diff to only emit records with changed fields since last run
Supported
Member-only checkout pricing
Requires active user cart session and authenticated state
Partial
User purchase history
PII and account-gated data is strictly excluded from all extraction pipelines
Partial
Infrastructure

Infrastructure powering the Zulily pipeline

Open-source tooling on proven cloud infra — no vendor lock-in, full observability.

ScrapyPlaywrightPython 3.12RedisPostgreSQLApache AirflowAWS LambdaS3CloudWatch2CaptchaCapSolverResidential ProxiesDockerKubernetesGrafanaPrometheus
Scrapy + Playwright Stack

Scrapy handles crawl orchestration and deduplication. Playwright handles JavaScript rendering, cookie sessions, and interaction flows for Zulily's dynamic frontend.

Residential Proxy Infrastructure

We maintain pools of residential ISP proxies across US regions. Rotation happens per-request with sticky sessions to maintain authenticated states.

Cloud-Native Orchestration

Pipelines run on Kubernetes for burst scaling during 6 AM deal launches. Airflow handles scheduling, dependency management, and SLA alerting.

Output & Delivery

Your data, your destination

Data delivered to where your team already works — no new tooling required.

JSON
Newline-delimited or nested array formats
CSV
Flat file with typed columns for spreadsheet analysis
XLS
Excel compatible format for business teams
Parquet
Columnar format optimised for big data analytics
AWS S3
Direct bucket delivery compatible with any data lake
Webhook
HTTP POST per record for real-time downstream processing
API
REST endpoints to query extracted datasets
Snowflake
Stage and COPY INTO workflow for incremental updates
S3
Direct bucket delivery — compatible with any data lake
// faq

Common questions.

About zulily.com scraping, legality, and pipeline operations.

Ask us directly →
Is scraping Zulily legal?

Scraping publicly available information from Zulily is generally permissible under applicable law. DataFlirt targets only public product, pricing, and event data. We do not extract personal data or user purchase histories. Clients should review Zulily's ToS and consult legal counsel for specific use cases.

How do you handle Zulily's login walls?

We manage authenticated cookie pools and use residential ISP proxies with realistic browser fingerprints. This allows us to bypass aggressive login prompts and access the public catalogue reliably.

Can you capture data right when flash sales launch?

Yes. We configure burst-scaling infrastructure to trigger extractions precisely at Zulily's daily launch times, ensuring you capture inventory and pricing before items sell out.

Do you extract all size and colour variations?

Yes. Our parsers map the complete variant matrix for every apparel item, capturing specific pricing, stock status, and waitlist availability for each size and colour combination.

Can you track waitlist metrics?

Yes. We extract the waitlist availability flags on sold-out items, providing valuable signals for product demand and inventory analysis.

What delivery formats do you support?

We deliver data in JSON, CSV, XLS, and Parquet formats. We can push directly to AWS S3, Snowflake, or trigger Webhooks for real-time integration.

Can I request a sample dataset?

Absolutely. We provide a sample run of up to 500 products as part of the pre-engagement scoping process so you can validate schema fit and data quality.

$ dataflirt scope --new-project --source=zulily.com ready

Tell us what
to extract.
We do the rest.

20-minute scoping call. Pilot dataset within the week. Production within two. Whether you need a one-off apparel catalogue dump or a continuous daily deal monitoring feed, we scope, build, and operate the pipeline. Tell us what you need.

hello@dataflirt.com · Bengaluru · IST · typical reply < 4h
Related Scrapers

More in fashion and apparel

Services

Data Extraction for Every Industry

View All Services →