
Diffbot
Structured web data platform that crawls the public web and exposes automatic extraction, knowledge graph, and search via API with free and paid plans.
Overview
Diffbot provides APIs that turn messy web pages into structured entities and relations without writing scrapers. Its automatic extractors read products, articles, discussions, and more, while the Knowledge Graph aggregates billions of facts organizations, people, and products for search and enrichment. Teams use it to power lead generation, market maps, and competitive tracking without maintaining parsers.
Pricing includes a free tier for testing and subscription plans that scale credits and endpoints, with higher tiers adding bulk extraction and advanced graph access. Developers integrate via REST with SDKs, then monitor usage in dashboards. Compared to building crawlers, Diffbot reduces maintenance and improves coverage, especially for product catalogs and research monitors that need fresh, structured signals.
Key features
- Automatic article and product extraction without custom rules
- Bulk extract and crawl capabilities for large scale coverage
- Global Knowledge Graph search and enrich endpoints
- Entity resolution and deduplication across sources
- REST APIs with SDKs and usage dashboards
- Flexible credits model that scales with volume
- Webhooks and batch jobs for reliable pipelines
- Support for research and lead enrichment workflows
Best for
- Enrich company and contact records for sales intelligence
- Track competitor product pages for price and spec changes
- Build vertical search on top of structured entities
- Assemble market maps by querying the Knowledge Graph
- Monitor news and blogs for emerging topics at scale
- Power research portals with fact level search
- Automate catalog ingestion for marketplaces
- Prototype analysts’ dashboards without scrapers
Capabilities
Auto parsers
Pull articles products discussions and more without writing site specific rules.
Entity search
Query a global knowledge graph to enrich companies people and products for analysis.
Crawl and Bulk
Run large jobs with bulk extract crawl scheduling and webhooks for robust pipelines.
Credits and SDKs
Manage usage with credits dashboards SDKs and alerts to keep costs predictable.
Frequently Asked Questions
How does pricing start?
Public plans show a free tier for testing and paid tiers starting near $299 per month with higher Plus plans at $899 per month.
Is there a free plan?
Yes, a free tier exists for basic testing and onboarding with limited credits.
Can I use my own crawler?
You can ingest URLs or rely on Diffbot crawl scheduling depending on your pipeline.
How fresh is the data?
Crawls and the Knowledge Graph update continuously with priority options on higher tiers.
Is enterprise SSO available?
Enterprise contracts add security features and support for governance needs.



