We extract factory profiles, material compositions, daily yarn pricing, and trade leads from Texnet. Delivered as clean JSON, CSV, or Parquet to S3, BigQuery, or Snowflake on your cadence.
Structured, schema-consistent data across all major object types — delivered clean, typed, and ready to query.
Complete list of extractable fields for Fabric & Yarn Listings objects from texnet.com.cn. All fields typed and schema-versioned.
"product_id": "TX-993821", "title": "100% Cotton Woven Poplin Fabric", "material_composition": "100% Cotton", "weight_gsm": 120, "width": "57/58 inches", "moq": 1000, "price_range_cny": "12.50 - 15.00", "yarn_count": "40s*40s"
| # | product_id | title | category | material_composition | weight_gsm | width |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Supplier Profiles objects from texnet.com.cn. All fields typed and schema-versioned.
"supplier_id": "SUP-44920", "company_name": "Shaoxing Keqiao Textile Co., Ltd.", "business_type": "Manufacturer, Trading Company", "location_province": "Zhejiang", "location_city": "Shaoxing", "established_year": 2008, "main_products": "['Polyester Fabric', 'Chiffon', 'Satin']", "verification_status": "Verified Member"
| # | supplier_id | company_name | business_type | location_province | location_city | established_year |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Trade Leads objects from texnet.com.cn. All fields typed and schema-versioned.
"lead_id": "TL-88392", "lead_type": "BUY", "product_category": "Spandex Yarn", "quantity_required": 5000, "unit_of_measure": "Kilograms", "posting_date": "2026-04-12", "buyer_location": "Guangdong", "status": "Active"
| # | lead_id | lead_type | product_category | quantity_required | unit_of_measure | posting_date |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Market Price Indices objects from texnet.com.cn. All fields typed and schema-versioned.
"index_id": "IDX-PTA-01", "material_type": "PTA (Purified Terephthalic Acid)", "region_market": "East China", "price_cny": 5820.0, "price_unit": "Ton", "date": "2026-05-10", "trend_percentage": 1.2, "specification": "Premium Grade"
| # | index_id | material_type | region_market | price_cny | price_unit | date |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Complete list of extractable fields for Industry News & Exhibitions objects from texnet.com.cn. All fields typed and schema-versioned.
"article_id": "NW-10293", "title": "Cotton Futures Surge Amidst Supply Chain Constraints", "publish_date": "2026-05-08", "category": "Market Analysis", "source": "Texnet Editorial", "tags": "['Cotton', 'Futures', 'Supply Chain', 'Pricing']", "author": "Li Wei"
| # | article_id | title | publish_date | author | category | content_body |
|---|---|---|---|---|---|---|
| 1 | ||||||
| 2 | ||||||
| 3 |
Our Texnet scraper navigates regional IP blocks, Mandarin localisation, and fragmented supplier catalogues to deliver clean, normalised procurement data.
Extract material composition, GSM, width, yarn count, and weaving techniques across millions of fabric listings.
Map factory locations, establishment years, registered capital, and ISO certifications for due diligence.
Track daily spot prices for raw materials like cotton, PTA, polyester, and viscose across regional Chinese markets.
Capture active buy and sell requests, including required quantities, delivery timelines, and buyer locations.
Automated parsing and standardisation of Chinese technical textile terminology into structured English schemas.
Parse phone numbers and email addresses hidden behind image-rendered text on supplier storefronts.
Utilise Mainland China residential proxy networks to bypass geo-blocking and access domestic market data.
Extract Minimum Order Quantities and volume-based pricing tiers to optimise procurement strategies.
Run daily or weekly pipelines to detect new supplier registrations, updated fabric catalogues, and price shifts.
Brief in. Clean data out.
Specify material categories, supplier regions, or price indices. We design the extraction schema together.
We configure Scrapy / Playwright crawlers, Mainland China proxy rotation, and OCR handling for texnet.com.cn.
Schema validation, Mandarin-to-English field verification, and sample supplier data before full launch.
JSON / CSV / Parquet pushed to your S3 bucket, BigQuery dataset, or Snowflake stage on agreed cadence.
Scraping Chinese B2B portals requires specialised infrastructure. Here is how we maintain data flow.
Texnet restricts or throttles traffic originating outside China. We route all requests through high-reputation residential ISP proxies physically located in Mainland China, ensuring uninterrupted access to domestic market data.
To prevent scraping, Texnet often renders supplier phone numbers and email addresses as images rather than text. Our pipeline includes an OCR step that reads these images and converts them back into structured text fields.
Textile specifications rely on highly specific Chinese terminology. We map and normalise standard industry terms (e.g., specific weave types, chemical compositions) into consistent English data types for your warehouse.
Texnet category pages often cap pagination at 100 pages, hiding deeper catalogue items. We programmatically segment searches by highly specific parameters (e.g., date ranges, micro-regions) to force the platform to expose the entire dataset.
Supplier microsites on Texnet frequently use JavaScript to load product grids and certification documents. We execute full Playwright browser sessions to hydrate the DOM before extraction.
Apparel brands and procurement teams identify alternative factories, verify certifications, and map supplier concentration in specific Chinese provinces.
Commodity analysts track daily spot prices for cotton, yarn, and synthetic fibres to forecast manufacturing costs and negotiate better terms.
Textile manufacturers monitor rival supplier catalogues, new fabric launches, and pricing adjustments to maintain market positioning.
Chemical suppliers and logistics companies extract active trade leads and factory contact details to build targeted outbound sales lists.
Fashion forecasting agencies analyse aggregate fabric listings to identify shifts in material composition preferences (e.g., rising demand for recycled polyester).
Audit firms scrape supplier profiles to cross-reference claimed ISO certifications and environmental compliance records against public registries.
"Texnet holds the ground truth for Chinese textile manufacturing — but extracting it requires navigating regional firewalls, technical Mandarin, and aggressive contact obfuscation."
Procurement teams waste hundreds of hours manually checking Texnet for price indices and factory details. DataFlirt automates this entirely. We handle the Mainland proxies, the OCR for hidden phone numbers, and the daily diffs, delivering a clean, normalised feed of the Chinese textile market directly to your infrastructure.
Everything supported by our texnet.com.cn scraper — rendered SPA elements, auth walls, rate-limit evasion and beyond.
Open-source tooling on proven cloud infra — no vendor lock-in, full observability.
Scrapy handles broad catalogue crawling and deduplication. Playwright is injected for specific supplier storefronts that require JavaScript execution to render product grids.
We maintain dedicated pools of residential ISP proxies within China. Rotation happens per-request to prevent IP bans and ensure consistent access to domestic-only pages.
Pipelines run on Kubernetes clusters. Airflow manages the complex dependency chains between index scraping, category discovery, and deep product extraction.
Data delivered to where your team already works — no new tooling required.
About texnet.com.cn scraping, legality, and pipeline operations.
Ask us directly →Scraping publicly available information from B2B directories is generally permissible for internal business intelligence. DataFlirt targets only public, non-authenticated supplier, product, and pricing data. We do not bypass login walls to extract VIP contact data. Clients should review platform terms and consult legal counsel for specific commercial use cases.
We route all extraction traffic through high-quality residential ISP proxies located in Mainland China. This ensures our requests appear as legitimate domestic traffic, bypassing regional blocks and throttling.
Yes. Texnet frequently uses image rendering to hide phone numbers and emails. Our pipeline includes an automated OCR (Optical Character Recognition) step that reads these images and converts them into structured text strings.
We normalise standard technical fields (like material composition, province names, and category structures) into English based on predefined mapping dictionaries. Free-text descriptions remain in the source language unless a specific translation pipeline is requested.
We can schedule the price index pipeline to run daily, capturing the latest spot prices for raw materials as soon as Texnet publishes them. You receive a timestamped diff of the daily changes.
Yes. When a category contains more results than the UI allows to be paginated, our crawlers automatically segment the search using micro-filters (e.g., narrowing by specific cities or price bands) to extract the complete catalogue.
Our smallest packages start at a defined extraction scope (e.g., all suppliers in Zhejiang province, or a specific fabric category) with weekly delivery. For full-site replication, we price based on volume and delivery frequency.
20-minute scoping call. Pilot dataset within the week. Production within two. Whether you need a one-off supplier directory export or a continuous daily feed of yarn pricing indices — we scope, build, and operate the pipeline. Tell us what you need.