Cookbook·Food delivery·papajohns.co.uk

Scrape listings on papajohns.co.uk

The call that returns listings from papajohns.co.uk today, with the parameters, patterns and failure modes read off the request log since 14 Aug 2026. Last tested 6 Sep 2026.

98.7% success · Aug 2026Browser neededStandard price · 1 credit

About papajohns.co.uk

papajohns.co.uk describes itself as “Papa Johns International has restaurants in over 50 countries. Use our International Restaurant Locator to eat Papa Johns Pizza all over the world”.

Its sitemaps list 503 URLs across 1 file. 445 of the 503 dated entries changed in the last 30 days, the newest on 6 Sep 2026.

Its robots.txt names 1 sitemap and disallows 5 paths, among them /order.

Measured on Crawlbase

Success rate over every request from every account in August 2026: 98.7%. Patterns, countries and failures come from the request log since 14 Aug 2026; 99.8% of successful calls used the JavaScript token and the median answer took 8.9s.

The call

Product pages are the shape the accounts fetch most here, with the country and token that succeed most often. Replace the placeholder path with a real product URL from the site.

Run in Playground
curl "https://api.crawlbase.com/?token=YOUR_JS_TOKEN&country=US&url=https%3A%2F%2Fwww.papajohns.co.uk%2Fjobs%2Fexample-section%2Fexample-page"

Parameters that matter

ParameterSet it toWhy
countryUS97.7% of successful calls set it and they succeed at 99.1%; calls without a country succeed at 100%.
javascriptuse the JavaScript token99.8% of calls use the JavaScript token and succeed at 99.2%. The plain token was tried too rarely (4 calls) to say anything about it.

URL patterns accounts fetch

PatternShare of successesJavaScript tokenCountryQuery parameters
/jobs/{slug}/{slug}56.5%100%USnone
/company/{slug}/{slug}7.0%100%USnone
/jobs/{slug}6.5%100%USradius
/company/{slug}5.0%100%USnone
/united-states/{slug}3.5%100%nonenone
/3.5%85.7%nonenone
/canada3.0%100%nonenone
/united-states3.0%100%nonenone

Fields you get

Structured data Crawlbase found on papajohns.co.uk, by page type. These are the schema.org types in the page, the fields your parser can read straight from the JSON-LD block:

  • Home: WebSite
  • Listing: WebSite, Organization
  • Article: WebSite, Organization
  • Other: LocalBusiness, WebSite, Organization

What breaks, and the fix

Of the failures since 14 Aug 2026, the site answered:

999: 19.0% of failures 999 means the site refused the fetch. Treat it like a block: retry once with the JavaScript token and the country the successful calls use, then stop.
504: 9.5% of failures A 504 is a timeout: the page did not finish inside the limit. Retry once, and keep your own client timeout above the median answer time shown for this site.
591: 61.9% of failures A 5xx is the site failing, not the fetch. Retry after a minute.

Run it on a schedule

A catalogue refresh is a list of product URLs. Push it to the Crawler with a callback, let it pace the calls and retry what the site refuses, and read the results from Cloud Storage. The median answer on this site is 8.9s.

Questions

Do I need a browser to scrape papajohns.co.uk?

Yes for most pages. JavaScript-token calls succeed at 99.2% while plain-token calls reach 100%.

How much does a papajohns.co.uk page cost on Crawlbase?

1 credit without a browser. papajohns.co.uk is in the standard tier.

Which country should I set for papajohns.co.uk?

US: that is what 97.7% of successful calls use.

Other pages on papajohns.co.uk

Pages papajohns.co.uk lists in its own sitemaps, fetched by Crawlbase on 6 Sep 2026 and verified in 2 consecutive runs. Each row is the cheapest call that returned the page with its content.

PageExampleTokenCountryMedian
Home
/
Structured data: WebSite
papajohns.co.uk/Plainnone7.0s
Listing
/blog/category/news
Structured data: WebSite, Organization
papajohns.co.uk/blog/category/newsPlainnone2.2s
Article
/blog/{slug}
Structured data: WebSite, Organization
papajohns.co.uk/blog/our-guide-to-pets-and-pizzaPlainnone1.9s
Other
/{slug}/{slug}
Structured data: LocalBusiness, WebSite, Organization
noneJavaScriptnone26.1s
Site names identify the pages a recipe fetches. Crawlbase is not affiliated with any site listed. Use recipes within the acceptable use policy.Measured numbers from August 2026, refreshed monthly