Scrape category pages on lego.com
The call that returns category pages from lego.com today, with the parameters, patterns and failure modes read off the request log since 14 Aug 2026. Last tested 6 Sep 2026.
About lego.com
lego.com describes itself as “Explore the world of LEGO® through games, videos, products and more! Shop awesome LEGO® building toys and brick sets and find the perfect gift for your kid”.
Its robots.txt names 6 sitemaps and disallows 12 paths, among them /identity/, /checkout, /innovation-intake-ideas.
Measured on Crawlbase
Success rate over every request from every account in August 2026: 94.9%. Patterns, countries and failures come from the request log since 14 Aug 2026; 0.0% of successful calls used the JavaScript token and the median answer took 2.8s.
Among the 20 ecommerce sites in the Cookbook, lego.com ranks 12 of 20 by success rate; the category median is 96.3%. Its median answer of 2.8s sits at the category median of 2.9s. 9 of the 20 accept the plain token, this one among them.
The call
Category and listing pages are the shape the accounts fetch most here, with the country and token that succeed most often. Replace the placeholder path with a real section URL from the site.
curl "https://api.crawlbase.com/?token=YOUR_TOKEN&country=DE&url=https%3A%2F%2Fwww.lego.com%2Fen-us%2Fcategories%2Fexclusives"from crawlbase import CrawlingAPI api = CrawlingAPI({'token': 'YOUR_TOKEN'}) r = api.get('https://www.lego.com/en-us/categories/exclusives', {'country': 'DE'}) html = r['body']
const { CrawlingAPI } = require('crawlbase'); const api = new CrawlingAPI({ token: 'YOUR_TOKEN' }); const r = await api.get('https://www.lego.com/en-us/categories/exclusives', { country: 'DE' });
Parameters that matter
| Parameter | Set it to | Why |
|---|---|---|
country | DE | 95.3% of successful calls set it and they succeed at 95.6%; calls without a country succeed at 83.7%. |
javascript | leave unset | Plain-token calls succeed at 95.4% and make up 99.6% of the traffic. A browser costs more and, on this site, buys nothing. |
URL patterns accounts fetch
| Pattern | Share of successes | JavaScript token | Country | Query parameters |
|---|---|---|---|---|
/de-de/categories/{slug} | 98.3% | 0.0% | DE | sort.key, sort.direction, page |
Fields you get
Structured data Crawlbase found on lego.com, by page type. These are the schema.org types in the page, the fields your parser can read straight from the JSON-LD block:
- Home:
Organization, WebPage - Listing:
ItemList, BreadcrumbList - Other:
ItemList, BreadcrumbList
What breaks, and the fix
Of the failures since 14 Aug 2026, the site answered:
Run it on a schedule
Walk the listings on a cadence to find new items, then push the item URLs to the Crawler with a callback; it paces the calls and retries what the site refuses. The median answer on this site is 2.8s.
Questions
Do I need a browser to scrape lego.com?
No. 95.4% of plain-token calls succeed and accounts send 99.6% of their traffic without a browser.
How much does a lego.com page cost on Crawlbase?
1 credit without a browser. lego.com is in the standard tier.
Which country should I set for lego.com?
DE: that is what 95.3% of successful calls use.
Other pages on lego.com
Pages lego.com lists in its own sitemaps, fetched by Crawlbase on 6 Sep 2026 and verified in 2 consecutive runs. Each row is the cheapest call that returned the page with its content.
| Page | Example | Token | Country | Median |
|---|---|---|---|---|
Home/Structured data: Organization, WebPage | lego.com/ | Plain | none | 3.5s |
Listing/en-us/categories/{slug}Structured data: ItemList, BreadcrumbList | lego.com/en-us/categories/exclusives | Plain | none | 5.9s |
Other/en-us/{slug}/{slug}Structured data: ItemList, BreadcrumbList | lego.com/en-us/themes/dc | Plain | none | 29.5s |