SYSTEM all green source cuyana.com queue 1,492 pages p99 latency 218ms dataflirt.com · scraper/cuyana-com
RUN · 14 active pipelines · cuyana.com live

Cuyana catalogue data,
at warehouse scale.

We extract premium bag listings, leather specifications, monogramming constraints, and inventory depth from Cuyana. Delivered as clean JSON, CSV, or Parquet to S3 or BigQuery on your cadence.

Products extracted
1,248 /run
Variant updates
6,410 /24h
Review records
42.8K /run
Active pipelines
14
Uptime
99.98%
Data Dictionary

Every field we extract from cuyana.com

Structured, schema-consistent data across all major object types — delivered clean, typed, and ready to query.

Complete list of extractable fields for Product Listings objects from cuyana.com. All fields typed and schema-versioned.

skutitlecategorysub_categorypricecurrencydescriptionmaterialsdimensionscare_instructionsimage_urlspage_url
product_listings
● 200 OK
"sku": "100111-001",
"title": "System Tote",
"price": 298.0,
"currency": "USD",
"category": "Bags",
"materials": "Italian Leather",
"dimensions": "11 in H x 19 in W x 5.5 in D"
# skutitlecategorysub_categorypricecurrency
1
2
3

Complete list of extractable fields for Variants & Colours objects from cuyana.com. All fields typed and schema-versioned.

parent_skuvariant_skucolour_namehex_codesizepricein_stockstock_levelbackorder_date
variants_& colours
● 200 OK
"parent_sku": "100111",
"variant_sku": "100111-001-OS",
"colour_name": "Black",
"in_stock": true,
"price": 298.0,
"size": "One Size"
# parent_skuvariant_skucolour_namehex_codesizeprice
1
2
3

Complete list of extractable fields for Monogramming Data objects from cuyana.com. All fields typed and schema-versioned.

skumonogram_eligiblemax_charactersfont_optionsfoil_coloursplacement_optionsmonogram_pricepreview_image_url
monogramming_data
● 200 OK
"sku": "100111-001",
"monogram_eligible": true,
"max_characters": 3,
"monogram_price": 15.0,
"foil_colours": "['Gold', 'Silver']",
"font_options": "['Serif', 'Sans Serif']"
# skumonogram_eligiblemax_charactersfont_optionsfoil_coloursplacement_options
1
2
3

Complete list of extractable fields for Reviews & Ratings objects from cuyana.com. All fields typed and schema-versioned.

review_idskureviewer_namestar_ratingreview_titlereview_bodyreview_dateverified_buyerlocation
reviews_& ratings
● 200 OK
"review_id": "REV-99281",
"sku": "100111-001",
"star_rating": 5,
"verified_buyer": true,
"review_date": "2023-11-04",
"review_title": "Perfect work bag"
# review_idskureviewer_namestar_ratingreview_titlereview_body
1
2
3

Complete list of extractable fields for Bundle & Add-On Data objects from cuyana.com. All fields typed and schema-versioned.

bundle_idbundle_titlebase_skuaddon_skusbundle_pricetotal_valuediscount_pctbundle_image_urlin_stock
bundle_& add-on data
● 200 OK
"bundle_id": "BNDL-TOTE-ORG",
"bundle_title": "System Tote + Insert",
"base_sku": "100111-001",
"bundle_price": 395.0,
"total_value": 425.0,
"in_stock": true
# bundle_idbundle_titlebase_skuaddon_skusbundle_pricetotal_value
1
2
3

Capabilities

Everything you need from Cuyana

Our scraper extracts deep catalogue data from Cuyana, handling dynamic variant loading, monogramming configuration logic, and high-resolution asset discovery.

Full Product Catalogue

Extract titles, descriptions, dimensions, care instructions, and material specifications for every item.

Colour & Variant Mapping

Map parent products to all colourways, capturing hex codes, specific pricing, and variant SKUs.

Monogram Configuration

Capture customisation rules including maximum characters, foil colours, and placement options per SKU.

Inventory Tracking

Monitor stock status, low stock warnings, and estimated backorder shipping dates.

High-Res Asset Extraction

Scrape full-resolution image URLs, lifestyle shots, and video assets for every product variant.

Bundle Pricing Logic

Extract 'System' bundle configurations, add-on pricing, and combined discount values.

Review & Sentiment Data

Paginate through customer reviews, capturing ratings, text, and verified buyer status.

International Pricing

Extract localised pricing and currency data based on target shipping destination.

Change Detection

Run continuous pipelines with hash-based diffing to track new product launches and colourway additions.

// engagement pipeline

From brand catalogue to structured data

Brief in. Clean data out.

Define Scope
d 0

Select target categories, specific product lines, or full catalogue extraction.

Pipeline Build
d 2–4

We configure Playwright crawlers to handle dynamic variant loading and monogramming previews.

Validation & QA
d 4–6

Schema validation, null-rate checks, and variant mapping verification before launch.

Delivery
ongoing

JSON, CSV, or Parquet pushed to your S3 bucket or warehouse on a defined schedule.

Under the hood

Handling direct-to-consumer platform complexities

Modern eCommerce frontends use heavy JavaScript and dynamic state. We handle the rendering layer.

pipeline-monitor · cuyana.com · live ● active
// fingerprinting
Identity rotation
TLS fingerprintrandomised
User-agentrotated
IP poolresidential
Challenges blocked0
// pagination
Page coverage
48,291 pages queued running
// observability
Pipeline health
99.9%
uptime
142ms
p99 lat
0.3%
null rate
2
alerts
JavaScript rendering
Playwright for dynamic variants

Cuyana loads variant specific data and pricing via frontend JavaScript. We use full browser rendering to capture accurate state for every colourway.

State management
Monogramming configuration logic

Customisation rules require specific user flows. Our crawlers simulate these flows to extract valid monogramming constraints and pricing.

Asset discovery
High-resolution image extraction

Product galleries use lazy loading and responsive image sources. We parse the underlying data structures to extract the highest resolution assets.

Inventory monitoring
Real-time stock signals

We monitor specific API endpoints and DOM elements to capture real-time inventory status, backorder dates, and low-stock indicators.

Anti-bot layer
Residential proxy rotation

We route requests through ISP-grade residential proxies to avoid rate limiting during high-frequency inventory checks.

Applications

Who uses Cuyana data

Teams across industries use cuyana.com data to build competitive products and smarter operations.

01
Competitor Benchmarking

Premium fashion brands monitor Cuyana's pricing architecture, material choices, and category expansion.

02
Assortment Planning

Retail strategists analyse colourway depth, bundle strategies, and product lifecycle duration.

03
Trend Forecasting

Fashion analysts track the introduction and retirement of specific leather finishes and silhouettes.

04
Market Research

Analyse customer reviews to identify common praise or complaints regarding specific materials or hardware.

05
Pricing Strategy

Track direct-to-consumer pricing models, bundle discounts, and international price localisation.

06
Supply Chain Analysis

Correlate out-of-stock events and backorder timelines to estimate production cycles and demand.

Why DataFlirt

"Understanding a premium direct-to-consumer brand requires tracking not just their products, but their exact inventory depth, colourway lifecycle, and bundle architecture."

Extracting data from modern headless eCommerce platforms requires executing JavaScript and understanding complex variant state. DataFlirt manages the rendering layer, proxy rotation, and schema normalisation so you receive clean, analysis-ready records without maintaining custom scraping infrastructure.

Technical Spec

Cuyana scraper — technical capabilities

Everything supported by our cuyana.com scraper — rendered SPA elements, auth walls, rate-limit evasion and beyond.

JavaScript rendering
Full Playwright sessions for variant and image gallery loading
Supported
Residential proxies
ISP-grade proxies to bypass rate limits during full catalogue crawls
Supported
Variant mapping
Link parent products to all sizes and colourways
Supported
Monogram rules extraction
Capture customisation constraints and pricing per item
Supported
Bundle logic parsing
Extract 'System' bundle components and combined pricing
Supported
Review pagination
Extract all historical customer reviews per product
Supported
High-res asset URLs
Extract uncompressed image and video asset links
Supported
Change detection
Only emit records with modified fields since the last run
Supported
User account history
Extraction of past orders and return status from user accounts
Partial
Store credit balances
Scraping of user-specific gift card or store credit balances
Partial
Infrastructure

Infrastructure powering the pipeline

Open-source tooling on proven cloud infra — no vendor lock-in, full observability.

ScrapyPlaywrightPython 3.12RedisPostgreSQLApache AirflowAWS LambdaS3CloudWatch2CaptchaCapSolverResidential ProxiesDockerKubernetesGrafanaPrometheus
Playwright Rendering

Handles Cuyana's dynamic frontend, executing JavaScript to load all variant data, monogramming previews, and high-resolution image galleries.

Proxy Infrastructure

Utilises residential ISP proxies to distribute requests during deep catalogue crawls, preventing rate limits and IP bans.

Managed Orchestration

Runs on Kubernetes with Apache Airflow scheduling. We monitor pipeline health, schema drift, and data completeness in real time.

Output & Delivery

Your data, your destination

Data delivered to where your team already works — no new tooling required.

JSON
Nested schema ideal for variant and bundle data
CSV
Flat file format for quick spreadsheet analysis
XLS
Excel format for business stakeholders
Parquet
Columnar format optimized for BigQuery and Snowflake
AWS S3
Direct delivery to your cloud storage buckets
Webhook
HTTP POST for real-time inventory alerts
API
REST endpoint to query latest extracted catalogue data
PostgreSQL
Direct database insertion with upsert logic
S3
Direct bucket delivery — compatible with any data lake
// faq

Common questions.

About cuyana.com scraping, legality, and pipeline operations.

Ask us directly →
Is scraping Cuyana legal?

Scraping publicly available product and pricing data is generally permissible. DataFlirt extracts public catalogue information and does not bypass authentication to access private user data.

How do you handle dynamic colour variants?

We use Playwright to execute the frontend JavaScript, simulating user interactions to load and extract data for every available colourway and size combination.

Can you extract monogramming options?

Yes. We parse the customisation configuration for each eligible product, capturing maximum character limits, available foil colours, and associated costs.

How frequently can you update inventory status?

We can configure pipelines to check stock levels and backorder dates on specific SKUs at daily or sub-daily intervals depending on your requirements.

Do you extract product bundles?

Yes. We identify bundle configurations, extracting the base product, available add-ons, and the calculated bundle discount pricing.

What happens if Cuyana changes their website design?

Our managed service includes continuous schema monitoring. If a DOM change breaks extraction, our engineering team updates the selectors to restore the pipeline.

Can I get historical pricing data?

We begin tracking pricing history from the moment your pipeline is commissioned. We do not retroactively generate historical price data prior to pipeline creation.

$ dataflirt scope --new-project --source=cuyana.com ready

Tell us what
to extract.
We do the rest.

20-minute scoping call. Pilot dataset within the week. Production within two. Get structured product, pricing, and inventory data delivered directly to your warehouse. Tell us your extraction requirements.

hello@dataflirt.com · Bengaluru · IST · typical reply < 4h
Related Scrapers

More in bags and luggage

Services

Data Extraction for Every Industry

View All Services →