Web Data Scraping Benchmarks
Provider success rates, response times and metadata coverage across our web scraping benchmarks. See the methodology.
Leaderboard
Filter by tool type and must-have requirements. Sort any column.
# | Model | Index | E-commerce | Social media | Search engines | Video & streaming | Travel | Reviews | Food delivery | App stores | Real estate | Job postings | News & publishing | Developer & SaaS | Finance | Government & education |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| 1 | Bright Data All tool types | 86.7 | 81.4 | 70.4 | 60.6 | 98.6 | 98.9 | 78 | 94.3 | 99.8 | 88.5 | 90.6 | 97.3 | 96.3 | 91.5 | 91.8 |
| 2 | Oxylabs All tool types | 77.2 | 19.8 | 70.4 | 0 | 99.7 | 89.1 | 56.2 | 90.8 | 99.6 | 71.1 | 77.5 | - | - | - | - |
| 3 | Apify All tool types | 75.4 | 18 | - | - | 99.6 | 99.2 | - | 55.7 | - | - | - | - | - | - | - |
| 4 | Nimble All tool types | 74 | 71.7 | 64 | 0 | 97.7 | 19.3 | 52.4 | 84.9 | 99.7 | 86 | 69.6 | 99.6 | 99.5 | - | 97 |
| 5 | Zyte All tool types | 72.5 | 25.5 | - | 6 | 99.6 | 96.3 | 65 | 82.2 | 99.5 | 65.6 | 57.8 | 97.6 | 98.6 | - | 96.7 |
| 6 | Decodo All tool types | 69.5 | 39.6 | 70.1 | 0 | 99.5 | 83.2 | 35.9 | 90 | 99.7 | 74.8 | 77.4 | - | - | - | - |
Cost vs. performance
Web scraper cost across our 100 e-commerce domain benchmark, and web unblocker cost across 10,000 domains. Volume-adjusted per successful request.
Avg is the effective per-1k rate we paid in our benchmark, volume-adjusted with each provider's cheapest available plan. Min is the cheapest tier each provider offers. Max is the most expensive tier.
Provider profile
Price per successful page, completion time, metadata depth, domains returning structured JSON and AI readiness for each provider.
# | Model | Domains with JSON | Avg metadata fields | $ per 1k | Completion time (s) | AI ready |
|---|---|---|---|---|---|---|
| 1 | Bright Data All tool types | 24 | 28 | 1.5 | 18 | 79 |
| 2 | Oxylabs All tool types | 12 | 4 | 2 | 14.2 | 71 |
| 3 | Apify All tool types | 10 | 42 | 13.7 | 26.7 | 55 |
| 4 | Nimble All tool types | 6 | 0 | 1 | 12.6 | 68 |
| 5 | Zyte All tool types | 0 | 0 | 3 | 13.6 | 60 |
| 6 | Decodo All tool types | 6 | 0 | 2.4 | 16.5 | 64 |
| 7 | SerpApi All tool types | 13 | 28 | - | 1.1 | - |
Behaviour under load
Success rate tracked at 1, 100, 500 and 5,000 parallel requests.
Anti-bot landscape
Which vendor guards your targets predicts your success rate.
Web unblocks across the top 10k domains
Scraping APIs across the top 10k domains
Domain coverage by provider
Which of the Tranco top 100 domains each provider can scrape, and for which of them it also returns structured JSON. Scraping support comes from our web unblocker benchmark; JSON support comes from a separate test of parsed output on the same domains. Domains are sorted by coverage, so the ones most providers support appear at the top.
| Domain | |||||||
|---|---|---|---|---|---|---|---|
| google.com | |||||||
| chatgpt.com | |||||||
| bing.com | |||||||
| amazon.com | |||||||
| facebook.com | |||||||
| tiktok.com | |||||||
| youtube.com | |||||||
| baidu.com | |||||||
| googlevideo.com | |||||||
| yandex.ru | |||||||
| apple.com | |||||||
| github.com | |||||||
| instagram.com | |||||||
| microsoft.com | |||||||
| office.com | |||||||
| twitter.com | |||||||
| wikipedia.org | |||||||
| yahoo.com | |||||||
| youtu.be | |||||||
| zoom.us | |||||||
| Domains scraped | 68 | 12 | 6 | 73 | 67 | 10 | 13 |
✓ The provider scrapes the domain. ✅ The provider scrapes the domain and returns structured JSON. ✕ The provider does not scrape the domain. Of the providers in this table, only Bright Data, Nimble and Zyte ran in the web unblocker benchmark; for the other providers the table shows JSON support only.
Web data providers markdown output performance
Markdown extraction success rate and completion time from our web unblocker benchmark, and the output formats each provider returns.
| Provider | Markdown | HTML |
|---|---|---|
Success rate by page type and concurrency
Product pages and search pages behave differently on the same domain.
Frequently asked questions
Methodology
The index re-uses the raw results of every web data benchmark we published. A benchmark counts if it tests web data providers and records the outcome of each request. A provider's index score is the average of its success rates across the benchmarks it ran in. Every benchmark carries equal weight, so a 1,400-request study counts as much as a 260,000-request one. Success means the response returned the target page with the content we asked for. We check that against bot pages and empty shells, so a 200 response with a challenge page counts as a failure. If a provider was never tested in a benchmark, we show a dash instead of a zero.
Targets come from the Tranco list, which ranks domains by averaging several traffic rankings. We remove dead, adult, gambling and malicious hosts before testing. Domains are grouped into categories: e-commerce, social media, search engines, video and streaming, travel, reviews, food delivery, app stores, real estate, job postings, news and publishing, developer and SaaS, finance, and government and education. Domain coverage and structured JSON support are checked on the top 100. Cost is per 1,000 successful pages, not per 1,000 requests sent. We take the cheapest plan a provider offers at that volume and divide by the success rate we measured. Completion time is how long one request takes from start to finish. Metadata fields are the parsed fields a provider returns when it answers with JSON. AI ready is the share of responses that came back as clean markdown. The anti-bot results come from two of our runs, the unblocker benchmark across 10,000 domains and the e-commerce benchmark across the top 100 domains, with each domain labelled by the vendor guarding it. Whiskers are 95% confidence intervals.
Explore Web Data Scraping Benchmarks
Best Twitter (X) Scrapers: Benchmarked
We benchmarked four web scraping APIs for Twitter scraping on the same 1,000 X post URLs, for 4,000 requests in total. For every request, we measured success rate and completion time. The fastest provider also had the highest success rate. Zyte completed 99.6% of the 1,000 X post requests at 8 seconds per request. Oxylabs…
Best Facebook Scrapers: Apify, Oxylabs & Decodo
We benchmarked five web scraping providers on the same 1,000 Facebook post URLs, for 5,000 requests in total. For every request, we measured success rate, completion time and the number of parsed metadata fields returned. See the best Facebook scraping tools based on supported page types, output formats, pricing, and trial options. Apify was the…
The Best LinkedIn Scrapers
We benchmarked four LinkedIn scraping APIs with 4,000 requests, sending each provider the same 1,000 LinkedIn post URLs. For every request we measured success rate, completion time, and the number of parsed metadata fields returned. Read our methodology for details about the LinkedIn benchmark. The chart below shows the daily success rate of each LinkedIn…
Best AI Web Scraping Tools Benchmarked
Sites change their layout and the fields you need from a page shift over time. These changes break manually-coded scrapers. AI scrapers can be updated with simple prompts, and some can repair a saved scraper when the site changes. We benchmarked top AI web scraping tools with 25,000 requests, 5,000 per tool, across 500 URLs…
Top 4 Google Play Scraping Providers Compared
We benchmarked four web scraping providers across Google Play product page URLs, sending 4,000 requests in total. For each request, we measured how reliably the provider returned data, how long it took from submission to final response, and how many metadata fields the response contained. Vendors are ranked by the number of structured metadata fields…
5 Best Scraping Browsers
Scraping browsers handle the unblocking infrastructure, enabling users to interact with websites programmatically and extract data easily. We benchmarked the top scraping browsers on sites with login walls, infinite scroll, and strict anti-bot rules. We updated this guide to include the latest anti-bot evasion techniques (TLS 1.3 fingerprinting) and updated pricing models for Bright Data…
We Tested the Best SERP Scraper APIs
We benchmarked 5 SERP providers using 7,000 live requests across Google, Bing, and Yandex. See which providers came out ahead on success rate and speed: Read the full methodology. We also run live requests every 15 min, with caching disabled, a 60s timeout, and 250+ queries across Google, Bing, and Yandex in the United States.…
Large-Scale Web Scraping: 7 Providers Benchmarked
We ran two benchmarks against live websites, from 5 to 5,000 concurrent requests. The first sent 260,000 requests through four web unblockers across the Tranco top 10,000 domains, plus a markdown extraction test on 10,000 URLs. The second fetched 65,000 product and search pages from each of five scraping providers across 100 e-commerce domains. Metrics…
Web Crawler Benchmark to Feed Websites to AI
We benchmarked four crawl APIs across three domains of varying difficulty at three max depth levels (5, 10, 20) with a 1,000-page limit, measuring crawl coverage, execution time, link discovery, markdown link quality, and title extraction accuracy. If you aim to: You can read our benchmark methodology. Firecrawl consistently crawled around 100 pages on theregister.com…
Web Scraping Roadmap
We scraped 10,000 live domains and 100 marketplaces using products from six web data infrastructure companies.We benchmarked these tools to see how well they handle enterprise web data use cases. We benchmarked 4 leading web data providers across the top 10,000 domains, running a total of 260,000 requests. Each provider was tested at multiple concurrency…