Cookbook·Reviews·yelp.com

Scrape biz pages on yelp.com

The call that returns biz pages from yelp.com today, with the parameters, patterns and failure modes read off the request log since 14 Aug 2026. Last tested 6 Sep 2026.

99.4% success · Aug 2026Browser on 51% of callsModerate price · 1.5 credits

About yelp.com

yelp.com describes itself as “User Reviews and Recommendations of Best Restaurants, Shopping, Nightlife, Food, Entertainment, Things to Do, Services and More at Yelp”.

Its robots.txt disallows 12 paths, among them /adtrack.

Measured on Crawlbase

Success rate over every request from every account in August 2026: 99.4%. Patterns, countries and failures come from the request log since 14 Aug 2026; 51.0% of successful calls used the JavaScript token and the median answer took 39.0s.

Among the 5 reviews sites in the Cookbook, yelp.com ranks 1 of 5 by success rate; the category median is 96.7%. Its median answer of 39.0s is slower than the category median of 16.3s. 1 of the 5 accept the plain token; this one does on some pages only.

The call

Article pages are the shape the accounts fetch most here, with the country and token that succeed most often. Replace the placeholder path with a real article URL from the site.

Run in Playground
curl "https://api.crawlbase.com/?token=YOUR_TOKEN&url=https%3A%2F%2Fwww.yelp.com%2Fbiz%2Fexample-page"

Parameters that matter

ParameterSet it toWhy
countryleave unsetMost successful calls set no country (100%) and succeed at 99.5%.
javascriptleave unsetPlain-token calls succeed at 99.5%, JavaScript-token calls at 99.4%. Both work; start without the browser and switch per URL pattern where the table below shows it.

URL patterns accounts fetch

PatternShare of successesJavaScript tokenCountryQuery parameters
/biz/{slug}50.5%100%nonestart
/{slug}/{slug}48.5%0.0%nonenone

Fields you get

Structured data Crawlbase found on yelp.com, by page type. These are the schema.org types in the page, the fields your parser can read straight from the JSON-LD block:

  • Detail page: WebPage, BreadcrumbList, WebSite, Organization, FAQPage

What breaks, and the fix

Of the failures since 14 Aug 2026, the site answered:

504: 52.7% of failures A 504 is a timeout: the page did not finish inside the limit. Retry once, and keep your own client timeout above the median answer time shown for this site.
400: 41.5% of failures 400 means the site refused the fetch. Treat it like a block: retry once with the JavaScript token and the country the successful calls use, then stop.

Run it on a schedule

News moves by the hour. Poll the section or the sitemap for new links, then push the article URLs to the Crawler with a callback so the fetches are paced. The median answer on this site is 39.0s.

Questions

Do I need a browser to scrape yelp.com?

It depends on the page. Plain-token calls succeed at 99.5%, JavaScript-token calls at 99.4%; the URL pattern table shows which paths accounts fetch with a browser.

How much does a yelp.com page cost on Crawlbase?

1.5 credits without a browser. yelp.com is in the moderate tier.

Which country should I set for yelp.com?

None is needed: most successful calls leave it unset.

Other pages on yelp.com

Pages yelp.com lists in its own sitemaps, fetched by Crawlbase on 6 Sep 2026 and verified in 2 consecutive runs. Each row is the cheapest call that returned the page with its content.

PageExampleTokenCountryMedian
Detail page
/{slug}/products/{slug}
Structured data: WebPage, BreadcrumbList, WebSite, Organization, FAQPage
nonePlainnone13.6s
Site names identify the pages a recipe fetches. Crawlbase is not affiliated with any site listed. Use recipes within the acceptable use policy.Measured numbers from August 2026, refreshed monthly