Cookbook·Books and audio·goodreads.com

Scrape book pages on goodreads.com

The call that returns book pages from goodreads.com today, with the parameters, patterns and failure modes read off the request log since 14 Aug 2026. Last tested 6 Sep 2026.

99.4% success · Aug 2026Browser neededStandard price · 1 credit

About goodreads.com

goodreads.com describes itself as “Find and read more books you’ll love, and keep track of the books you want to read. Be part of the world’s largest community of book lovers on Goodreads”.

Its robots.txt names 16 sitemaps and disallows 12 paths, among them /admin, /api, /ebooks. It asks crawlers to wait 5s between requests.

Measured on Crawlbase

Success rate over every request from every account in August 2026: 99.4%. Patterns, countries and failures come from the request log since 14 Aug 2026; 97.5% of successful calls used the JavaScript token and the median answer took 10.0s.

Among the 3 books and audio sites in the Cookbook, goodreads.com ranks 2 of 3 by success rate; the category median is 99.4%. Its median answer of 10.0s sits at the category median of 10.0s. 1 of the 3 accept the plain token; this one needs the JavaScript token.

The call

Product pages are the shape the accounts fetch most here, with the country and token that succeed most often. Replace the placeholder path with a real product URL from the site.

Run in Playground
curl "https://api.crawlbase.com/?token=YOUR_JS_TOKEN&url=https%3A%2F%2Fwww.goodreads.com%2Fbook%2Fexample-section%2Fexample-page"

Parameters that matter

ParameterSet it toWhy
countryleave unsetMost successful calls set no country (100%) and succeed at 99.6%.
javascriptuse the JavaScript token95.6% of calls use the JavaScript token and succeed at 99.6%. The plain token was tried too rarely (22 calls) to say anything about it.

URL patterns accounts fetch

PatternShare of successesJavaScript tokenCountryQuery parameters
/book/{slug}/{slug}49.0%100%nonefrom_choice
/author/{slug}/{slug}44.5%100%nonenone
/book/{slug}3.3%0.0%nonecategory, country, duration
/book/{slug}/{id}2.0%100%nonenone

Fields you get

A product page usually carries an ld+json Product block with name, price, currency and availability, and the same values in the visible markup. Read the block first; it changes less often than the layout.

What breaks, and the fix

Of the failures since 14 Aug 2026, the site answered:

504: 100% of failures A 504 is a timeout: the page did not finish inside the limit. Retry once, and keep your own client timeout above the median answer time shown for this site.

Run it on a schedule

A catalogue refresh is a list of product URLs. Push it to the Crawler with a callback, let it pace the calls and retry what the site refuses, and read the results from Cloud Storage. The median answer on this site is 10.0s.

Questions

Do I need a browser to scrape goodreads.com?

Yes for most pages. JavaScript-token calls succeed at 99.6% while plain-token calls reach 100%.

How much does a goodreads.com page cost on Crawlbase?

1 credit without a browser. goodreads.com is in the standard tier.

Which country should I set for goodreads.com?

None is needed: most successful calls leave it unset.

Other pages on goodreads.com

Pages goodreads.com lists in its own sitemaps, fetched by Crawlbase on 6 Sep 2026 and verified in 2 consecutive runs. Each row is the cheapest call that returned the page with its content.

PageExampleTokenCountryMedian
Home
/
goodreads.com/Plainnone2.6s
Profile
/author/{slug}/{slug}
goodreads.com/author/show/61105.Dr_SeussPlainnone2.0s
Site names identify the pages a recipe fetches. Crawlbase is not affiliated with any site listed. Use recipes within the acceptable use policy.Measured numbers from August 2026, refreshed monthly