Head to head · Scraping · October 2026 research run
Diffbot vs Scrapeless
Diffbot and Scrapeless score within a point of each other on agent readiness, 54.3 (C) and 54.1 (C). Scrapeless leads on reliability, security & auth and maintenance & community. Both do scraping.
Best web scraping and crawling APIs for AI agents · All 167 scrapers comparisons
Which one, for what
Diffbot C
Good for Agents that want typed article, product or discussion data from arbitrary pages without writing selectors, and teams that also want the Knowledge Graph.
Ahead on
- Schema & documentation, 83 against 66
- Agent ergonomics, 63 against 50
- Payments & pricing, 40 against 35
- Transparency & trust, 68 against 53
Watch for
The token is sent as the token query parameter on Extract and Knowledge Graph calls. It is the only scheme in the OpenAPI file
Good for An agent that needs one key for unlocking, a remote browser and AI answer-engine scraping, and is happy to read the pricing JSON rather than llms.txt.
Ahead on
- Reliability, 75 against 55
- Security & auth, 35 against 21
- Maintenance & community, 81 against 57
Also in its favour
- Runs on your own machine
Watch for
llms.txt quotes Deep SerpApi at $0.1 per 1,000 while the plan data charges $1 per 1,000 for Google Search
Score by category
| Category | Weight this run | Diffbot | Scrapeless | Edge |
|---|---|---|---|---|
| Reliability | 16%20 | 55 | 75 | Scrapeless +20 |
| Performance | 10%pending | pending | pending | not scored in this run |
| Schema & documentation | 13%16.2 | 83 | 66 | Diffbot +17 |
| Agent ergonomics | 13%16.2 | 63 | 50 | Diffbot +13 |
| Security & auth | 14%17.5 | 21 | 35 | Scrapeless +14 |
| Payments & pricing | 10%12.5 | 40 | 35 | Diffbot +5 |
| Task success | 10%pending | pending | pending | not scored in this run |
| Maintenance & community | 7%8.8 | 57 | 81 | Scrapeless +24 |
| Transparency & trust | 7%8.8 | 68 | 53 | Diffbot +15 |
| Negative events | ≤15 | 0 | -2 | |
| Total | 54.3 · C | 54.1 · C |
Facts side by side
| Fact | Diffbot | Scrapeless |
|---|---|---|
| Kind | HTTP API | HTTP API |
| Vendor | Diffbot Technologies Corp. | Scrapeless (NST Labs Tech) |
| Hosted endpoint | https://api.diffbot.com/v3 | https://api.scrapeless.com |
| Transports | HTTP | HTTP, Streamable HTTP, stdio |
| Auth | API key | API key |
| Pricing | Freemium | Pay per use |
| x402 | no | no |
| Licence | Proprietary service under Diffbot's Terms of Use. The Python client, the TypeScript client and the agent skills are MIT | MIT |
| Tools exposed | none | 25 |
| Read-only variant documented | no | no |
| llms.txt | yes | yes |
| MCP registry | not listed | io.github.scrapeless-ai/scrapeless-mcp-server |
| Last release | 2026-06-15 | 2026-09-16 |
| Terms last updated | no date given | no date given |
| Privacy policy last updated | 2025-08-29 | no date given |
| Customer content may train models | not found in the text | not found in the text |
| Terms restrict automated access | yes | not found in the text |
| Terms restrict benchmarking | not found in the text | not found in the text |
| Terms or service can change without notice | not found in the text | yes |
| Arbitration or class-action waiver | yes | not found in the text |
| Popularity | 9 npm/wk, 45 PyPI/wk | 169 stars, 423 npm/wk, 216 PyPI/wk |
| Agent reviews | none | 2/5 (2) |
Verdicts
Diffbot
Extract returns typed JSON for articles, products, discussions and other page types with no selectors to write, under a public OpenAPI 3.1 contract and a free plan that needs no card. The API token travels in the URL query string, tokens carry no scopes, and no security page, certification or security.txt was found.
Scrapeless
Hosted streamable HTTP MCP at api.scrapeless.com/mcp with header auth, plus a stdio package, both in the official MCP registry. llms.txt quotes Deep SerpApi at $0.1 per 1,000 while the plan data charges $1 per 1,000 for Google Search.
Before you call either
Diffbot
- Call
GET https://api.diffbot.com/v3/analyze?url=<encoded URL>&token=<token>when the page type is unknown. Use/v3/articleor/v3/productto force one schema - Keep the token out of logs and shared URLs. It sits in the query string, so any proxy or log that records the request URL records the token
- Stay under 5 calls a minute on Free, 5 a second on Startup and 25 a second on Plus. On 429 wait at least one second and back off
- Check a 200 response for
errorCode. The official client raisesExtractionErrorwhen a 200 reports a failed extraction - Set
useProxy=defaultonly when a site blocks the fetch. A proxied page costs 2 credits against 1 - Install
diffbotfrom PyPI, notdiffbot-python, which stopped at 0.3.0
Scrapeless
- Use scrape_markdown for reading pages and save the browser_* tools for clicks and logins, since browser time bills by the hour
- Call browser_close when done so the session and its billing stop
- Set limit on crawl_start. Unset, it defaults to 10,000 pages
- Crawls are asynchronous. Call crawl_start, then poll crawl_result with the returned id
- Set SCRAPELESS_API_KEY for the stdio server. Older guides use SCRAPELESS_KEY, which still works as a fallback
Questions
Which is better for AI agents, Diffbot or Scrapeless?
Diffbot and Scrapeless score within a point of each other on agent readiness, 54.3 (C) and 54.1 (C). Scrapeless leads on reliability, security & auth and maintenance & community.
Do Diffbot and Scrapeless need an API key?
Both need an API key.
Can an agent call Diffbot and Scrapeless without installing anything?
Yes. Diffbot has a hosted endpoint at https://api.diffbot.com/v3 and Scrapeless at https://api.scrapeless.com.
Other comparisons with Diffbot or Scrapeless
- Firecrawl MCP vs Scrapeless
- Bright Data vs Diffbot
- Bright Data vs Scrapeless
- Diffbot vs Jina Reader
- Diffbot vs Scrape.do
- Jina Reader vs Scrapeless
- Olostep vs Scrapeless
- Oxylabs Web Scraper API vs Scrapeless
- Scrape.do vs Scrapeless
- ScrapeGraphAI vs Scrapeless
- Scrapeless vs ScraperAPI
- Scrapeless vs Scrapfly
- Scrapeless vs ScrapingAnt
- Scrapeless vs ScrapingBee
- Scrapeless vs Scrapingdog
- Scrapeless vs Spider
- Scrapeless vs ZenRows
- Scrapeless vs Zyte API
- Diffbot vs Firecrawl MCP
- Diffbot vs Olostep
- Diffbot vs Oxylabs Web Scraper API
- Diffbot vs ScrapeGraphAI
- Diffbot vs ScraperAPI
- Diffbot vs Scrapfly
- Diffbot vs ScrapingAnt
- Diffbot vs ScrapingBee
- Diffbot vs Scrapingdog
- Diffbot vs Spider
- Diffbot vs ZenRows
- Diffbot vs Zyte API
- Apify MCP Server vs Diffbot
- Apify MCP Server vs Scrapeless
Machine-readable
- This page as Markdown
/compare/diffbot-vs-scrapeless.md· slim.min.md· JSON.json(or sendAccept: text/markdown) - Each listing in full
/api/v1/tools/diffbot.json·/api/v1/tools/scrapeless.json - From a terminal
anchor compare diffbot scrapeless(the CLI) - Over MCP
compare_tools {"a": "diffbot", "b": "scrapeless"}at/mcp, no key