Scrape product pages on shein.com
The call that returns product pages from shein.com today, with the parameters, patterns and failure modes read off the request log since 14 Aug 2026. Last tested 6 Sep 2026.
About shein.com
shein.com describes itself as “Free shipping on orders from 590 RUB.✓1000+ new arrivals daily✓100% safe online shopping✓In clothes from SHEIN you can become whoever you want! A large assortment of outfits and a variety of styles! Great promotions and sales! New outfits are added daily!”.
Its home page declares 12 languages: en-us, es-us, en, es, pt, de-de, ar-de, en-de, de-at, ar-at, tr-at, ro-at.
Its robots.txt names 1 sitemap and disallows 7 paths, among them /user/, /cart/, /geetest/.
Measured on Crawlbase
Success rate over every request from every account in August 2026: 98.1%. Patterns, countries and failures come from the request log since 14 Aug 2026; 0.0% of successful calls used the JavaScript token and the median answer took 14.0s.
Among the 20 fashion sites in the Cookbook, shein.com ranks 12 of 20 by success rate; the category median is 98.5%. Its median answer of 14.0s sits at the category median of 13.5s. 10 of the 20 accept the plain token, this one among them.
The call
Product pages are the shape the accounts fetch most here, with the country and token that succeed most often. Replace the placeholder path with a real product URL from the site.
curl "https://api.crawlbase.com/?token=YOUR_TOKEN&url=https%3A%2F%2Fwww.shein.com%2Fexample-section%2Fexample-page"from crawlbase import CrawlingAPI api = CrawlingAPI({'token': 'YOUR_TOKEN'}) r = api.get('https://www.shein.com/example-section/example-page') html = r['body']
const { CrawlingAPI } = require('crawlbase'); const api = new CrawlingAPI({ token: 'YOUR_TOKEN' }); const r = await api.get('https://www.shein.com/example-section/example-page');
Parameters that matter
| Parameter | Set it to | Why |
|---|---|---|
country | leave unset | Most successful calls set no country (94.4%) and succeed at 77.7%. |
javascript | leave unset | Plain-token calls succeed at 85.8% and make up 85.5% of the traffic. A browser costs more and, on this site, buys nothing. |
scraper | shein-product | 8.5% of calls request this scraper and get structured JSON back instead of HTML. |
URL patterns accounts fetch
| Pattern | Share of successes | JavaScript token | Country | Query parameters |
|---|---|---|---|---|
/{slug}/{slug} | 89.8% | 0.0% | none | src_identifier, sort, page |
/api/{slug}/{slug} | 10.3% | 0.0% | none | _ver, _lang, type |
Fields you get
Structured data Crawlbase found on shein.com, by page type. These are the schema.org types in the page, the fields your parser can read straight from the JSON-LD block:
- Home:
OnlineStore - Other:
BreadcrumbList, ItemList
What breaks, and the fix
Of the failures since 14 Aug 2026, the site answered:
Run it on a schedule
A catalogue refresh is a list of product URLs. Push it to the Crawler with a callback, let it pace the calls and retry what the site refuses, and read the results from Cloud Storage. The median answer on this site is 14.0s.
Questions
Do I need a browser to scrape shein.com?
No. 85.8% of plain-token calls succeed and accounts send 85.5% of their traffic without a browser.
How much does a shein.com page cost on Crawlbase?
2 credits without a browser. shein.com is in the complex tier.
Which country should I set for shein.com?
None is needed: most successful calls leave it unset.
Other pages on shein.com
Pages shein.com lists in its own sitemaps, fetched by Crawlbase on 6 Sep 2026 and verified in 2 consecutive runs. Each row is the cheapest call that returned the page with its content.
| Page | Example | Token | Country | Median |
|---|---|---|---|---|
Home/Structured data: OnlineStore | shein.com/ | Plain | none | 73.5s |
Other/{slug}/{slug}Structured data: BreadcrumbList, ItemList | none | Plain | none | 51.6s |