serp.fast
The web data tool directory

The directory for web scraping and AI data tools

83 search APIs, extraction tools, and browser platforms across 8 categories, reviewed for teams building AI products that read the web.

Nathan Kessler

Curated by Nathan Kessler · Read weekly by AI product builders

These sponsor slots reach founders and PMs evaluating web data tools.

Browse the directory

Embeddings-based neural search that finds semantically related pages

FreemiumMar 2026AI-Native Search APIs

Real-time search API built for AI agents and RAG pipelines

FreemiumMar 2026AI-Native Search APIs

LLM search API that returns cited answers instead of raw links

PaidMar 2026AI-Native Search APIs

Composable web, news, and research search APIs for enterprise agents

PaidMar 2026AI-Native Search APIs

Proprietary web index built for deep multi-hop agent research

FreemiumMar 2026AI-Native Search APIs

AI search that licenses publisher content instead of scraping it

FreemiumMar 2026AI-Native Search APIs

Reader API that turns any URL into clean markdown, now part of Elastic

FreemiumMar 2026AI-Native Search APIs

Model-agnostic web access you add to any LLM in one call

FreemiumMar 2026AI-Native Search APIs

Ad-free, privacy-first AI search that answers rather than lists links

FreemiumMar 2026AI-Native Search APIs

Self-hosted metasearch that aggregates results from 70+ engines

FreeMar 2026AI-Native Search APIs

Web knowledge API running its own Rust crawler and hybrid index

FreemiumJun 2026AI-Native Search APIs

Search API spanning the web plus arXiv, PubMed, and SEC filings

FreemiumJun 2026AI-Native Search APIs

Independent search index behind a single JSON endpoint for agents

FreemiumJun 2026AI-Native Search APIs

Recall-first news and web search across 140,000+ sources

FreemiumMay 2026AI-Native Search APIs

SERP scraping API covering 80+ search engines since 2017

PaidMar 2026SERP Data APIs

Fast Google SERP API priced at a dollar per thousand queries

FreemiumMar 2026SERP Data APIs

SEO data platform pairing SERP results with 50+ other APIs

PaidMar 2026SERP Data APIs

Multi-engine SERP API with native LangChain and Haystack support

FreemiumMar 2026SERP Data APIs

Low-cost SERP API with batch jobs up to 15,000 requests

FreemiumMar 2026SERP Data APIs

SERP API wired into Zapier and Make.com for no-code pipelines

FreemiumMar 2026SERP Data APIs

Volume-priced SERP API that drops toward 29 cents per thousand

FreemiumMar 2026SERP Data APIs

Proxy aggregator and scraper monitoring with a free tutorial hub

FreemiumMar 2026SERP Data APIs

Google SERP API with image and reverse image search built in

FreemiumMar 2026SERP Data APIs

Turns any website into LLM-ready markdown via a single API

Freemium155K+Web Crawl & Data Extraction APIs

Scraping platform with a marketplace of 50,000+ prebuilt Actors

FreemiumMar 2026Web Crawl & Data Extraction APIs

Enterprise extraction platform from the team behind Scrapy

PaidMar 2026Web Crawl & Data Extraction APIs

Handles proxy rotation and JS rendering behind one scraping API

FreemiumMar 2026Web Crawl & Data Extraction APIs

Gets scrapers past CAPTCHAs, fingerprinting, and WAF blocks

PaidMar 2026Web Crawl & Data Extraction APIs

Anti-detection scraping API with CAPTCHA solving at low cost

FreemiumMar 2026Web Crawl & Data Extraction APIs

Fully managed scraping service with SLAs and human QA

EnterpriseMar 2026Web Crawl & Data Extraction APIs

Bundles anti-bot bypass, rendering, and screenshots into one API

FreemiumMar 2026Web Crawl & Data Extraction APIs

Web data platform with native Databricks and Snowflake pipes

FreemiumMar 2026Web Crawl & Data Extraction APIs

Adaptive Python scraper whose selectors survive site redesigns

Free63K+Web Crawl & Data Extraction APIs

Extracts structured JSON with prebuilt, agent-ready site actions

FreemiumApr 2026Web Crawl & Data Extraction APIs

Rust page-to-markdown API with raw-HTTP-first browser-fallback engine

Freemium1.8K+Web Crawl & Data Extraction APIs

Mozilla's web execution API routing from raw fetch to full browser

FreemiumMay 2026Web Crawl & Data Extraction APIs

URL-to-markdown API with an added brand-data layer for agents

FreemiumJul 2026Web Crawl & Data Extraction APIs

Cloud browser infrastructure for agents, with the Stagehand SDK

FreemiumMar 2026Browser Infrastructure

Open-source browser API with stealth and session management

FreemiumMar 2026Browser Infrastructure

Headless browser platform for scraping, PDFs, and screenshots

Freemium13K+Browser Infrastructure

Browser automation for agents, and a verified Cloudflare bot

FreemiumMar 2026Browser Infrastructure

Browser-as-a-service with natural-language control and fast launches

FreemiumMar 2026Browser Infrastructure

Natural-language cloud browser with SOC2 and HIPAA compliance

FreemiumMar 2026Browser Infrastructure

Zig-built headless browser using far less memory than Chrome

Freemium31K+Browser Infrastructure

Autoscaling browsers with anti-bot detection and reusable sessions

FreemiumMar 2026Browser Infrastructure

Hands any website to an AI agent, the top browser-automation repo

Freemium98K+Browser Infrastructure

Anti-detect browser with fingerprint control and a headless API

FreemiumMar 2026Browser Infrastructure

Patches Chromium to run automation that bot detectors miss

FreemiumMar 2026Browser Infrastructure

API for the only independent Western search index at scale

FreemiumMar 2026Independent Web Indexes

Nonprofit web archive of over 10 petabytes behind most major LLMs

FreeMar 2026Independent Web Indexes

UK search engine with its own index and a no-tracking pledge

FreemiumMar 2026Independent Web Indexes

Structured feeds of news, forum, and dark-web content for enterprise

EnterpriseMar 2026Independent Web Indexes

Programmatic access to Yandex index, dominant across Russia

PaidMar 2026Independent Web Indexes

Search engine for the text-heavy old web Google buries

FreeMar 2026Independent Web Indexes

French GDPR-native search running its own partial index

PaidMar 2026Independent Web Indexes

Vision and NLP parsing behind a 10-billion-entity knowledge graph

PaidMar 2026Agentic Extraction

Python library that scrapes sites from plain-English prompts

Freemium27K+Agentic Extraction

TypeScript and Python SDK for driving browsers with natural language

Free23K+Agentic Extraction

Describe the data you need and get a REST API for any site

FreemiumMar 2026Agentic Extraction

Prompt-driven web research that builds and enriches spreadsheet rows

FreemiumMar 2026Agentic Extraction

Enterprise web data extraction that auto-generates logic, no code

FreemiumMar 2026Agentic Extraction

AI agent that navigates and extracts using vision and LLMs

Freemium21K+Agentic Extraction

Query language for pulling structured data from any web page

FreemiumMar 2026Agentic Extraction

Point-and-click robots that extract and monitor any site

FreemiumJun 2026Agentic Extraction

The original Python framework for large-scale web crawling

Free62K+Open Source Frameworks

Microsoft's cross-browser automation for testing and scraping

Free90K+Open Source Frameworks

Google's Node.js library for driving headless Chrome

Free94K+Open Source Frameworks

Apify's crawling library wrapping Playwright and Puppeteer

Free23K+Open Source Frameworks

Veteran browser automation with bindings for most languages

Free34K+Open Source Frameworks

Open-source crawler shaped for RAG, among the most-starred on GitHub

Free68K+Open Source Frameworks

Python parser that turns messy HTML into navigable trees

FreeMar 2026Open Source Frameworks

jQuery-style HTML parsing for Node.js, no browser required

Free30K+Open Source Frameworks

Fast concurrent scraping framework for Go with a callback API

Free25K+Open Source Frameworks

Pairs Requests and Beautiful Soup for form-driven scraping

Free4.8K+Open Source Frameworks

Lightweight async scraping from HTTPx plus Scrapy Parsel

FreeMar 2026Open Source Frameworks

Strips nav, ads, and chrome to leave the main article text

Free6.1K+Open Source Frameworks

The library behind Firefox Reader Mode, extracting article text

Free11K+Open Source Frameworks

Benchmark testing browser agents on 153 tasks across 144 live sites

FreeApr 2026Benchmarks

812 long-horizon web tasks on reproducible self-hosted sites

FreeMay 2026Benchmarks

643 live-web tasks measuring end-to-end multimodal agents

FreeMay 2026Benchmarks

2,350 tasks across 137 sites measuring cross-domain transfer

FreeMay 2026Benchmarks

369 computer-use tasks spanning Ubuntu, Windows, and macOS

FreeMay 2026Benchmarks

Scores how web search tools lift agent accuracy on research tasks

FreeJul 2026Benchmarks

Evaluating alternatives to...

View all alternatives

Bright Data is the largest traditional proxy and web data vendor, built around residential IP infrastructure rather than AI-native developer experience. Most teams who land here are evaluating whether the legacy enterprise contract is still the right fit for an AI product, or whether a newer scraping API can do the same job with cleaner integration and lower per-page cost.

7 alternatives

Oxylabs is the European enterprise scraping platform with a product matrix close to Bright Data's: Web Scraper API, SERP Scraper API, E-Commerce Scraper API, Web Unblocker, and residential and mobile proxies, plus a newer AI Studio suite (AI Scraper, AI Crawler, MCP server) that returns LLM-ready Markdown. Teams evaluating Oxylabs typically reach this page when the enterprise contract feels like overkill for what their AI product actually needs.

7 alternatives

Decodo rebranded from Smartproxy in April 2025 and now markets itself as an AI-ready proxy and scraping platform, but the foundation and pricing stay proxy-bandwidth-first. Its flagship residential product bills per GB, which is hard to map to an AI workload measured in pages or documents. Teams land here when the AI layer (AI Parser, MCP server, n8n node, Markdown output) turns out to be bolted onto proxy infrastructure rather than a purpose-built extraction API, and they want a single clean call that returns LLM-ready data.

7 alternatives

Teams that shortlist ScrapingBee are usually building an AI agent or pipeline that needs to fetch and read arbitrary web pages. ScrapingBee is a request-based scraping API that bundles proxies, headless-Chrome rendering, and anti-bot bypass behind one endpoint. It is a low-level HTML fetcher, but its credit-multiplier pricing and weak success rate on defended targets push many builders toward tools that return LLM-ready data directly. The alternatives below are all in our directory and lead with that AI-native posture.

7 alternatives