Embeddings-based neural search that finds semantically related pages
CategoryAI-Native Search APIs Real-time search API built for AI agents and RAG pipelines
CategoryAI-Native Search APIs LLM search API that returns cited answers instead of raw links
CategoryAI-Native Search APIs Composable web, news, and research search APIs for enterprise agents
CategoryAI-Native Search APIs Proprietary web index built for deep multi-hop agent research
CategoryAI-Native Search APIs AI search that licenses publisher content instead of scraping it
CategoryAI-Native Search APIs Reader API that turns any URL into clean markdown, now part of Elastic
CategoryAI-Native Search APIs Model-agnostic web search you add to any LLM in one call
CategoryAI-Native Search APIs Ad-free, privacy-first AI search that answers rather than lists links
CategoryAI-Native Search APIs Self-hosted metasearch that aggregates results from 70+ engines
CategoryAI-Native Search APIs Web knowledge API running its own Rust crawler and hybrid index
CategoryAI-Native Search APIs Search API spanning the web plus arXiv, PubMed, and SEC filings
CategoryAI-Native Search APIs Independent search index behind a single JSON endpoint for agents
CategoryAI-Native Search APIs Recall-first news and web search across 140,000+ sources
CategoryAI-Native Search APIs SERP scraping API covering 80+ search engines since 2017
Fast Google SERP API priced at a dollar per thousand queries
SEO data platform pairing SERP results with 50+ other APIs
Multi-engine SERP API with native LangChain and Haystack support
Low-cost SERP API with batch jobs up to 15,000 requests
SERP API wired into Zapier and Make.com for no-code pipelines
Volume-priced SERP API that drops toward 29 cents per thousand
Proxy aggregator and scraper monitoring with a SERP API attached
Google SERP API with image and reverse image search built in
Turns any website into LLM-ready markdown via a single API
CategoryWeb Crawl & Data Extraction APIs Scraping platform with a marketplace of 10,000+ prebuilt Actors
CategoryWeb Crawl & Data Extraction APIs Enterprise extraction platform from the team behind Scrapy
CategoryWeb Crawl & Data Extraction APIs Handles proxy rotation and JS rendering behind one scraping API
CategoryWeb Crawl & Data Extraction APIs Gets scrapers past CAPTCHAs, fingerprinting, and WAF blocks
CategoryWeb Crawl & Data Extraction APIs Anti-detection scraping API with CAPTCHA solving at low cost
CategoryWeb Crawl & Data Extraction APIs Fully managed scraping service with SLAs and human QA
CategoryWeb Crawl & Data Extraction APIs Bundles anti-bot bypass, rendering, and screenshots into one API
CategoryWeb Crawl & Data Extraction APIs Web data platform with native Databricks and Snowflake pipes
CategoryWeb Crawl & Data Extraction APIs Adaptive Python scraper whose selectors survive site redesigns
CategoryWeb Crawl & Data Extraction APIs Extracts structured JSON with prebuilt, agent-ready site actions
CategoryWeb Crawl & Data Extraction APIs Rust page-to-markdown API that skips the headless browser
CategoryWeb Crawl & Data Extraction APIs Mozilla's web execution API routing from raw fetch to full browser
CategoryWeb Crawl & Data Extraction APIs URL-to-markdown API with an added brand-data layer for agents
CategoryWeb Crawl & Data Extraction APIs Cloud browser infrastructure for agents, with the Stagehand SDK
CategoryBrowser Infrastructure Open-source browser API with stealth and session management
CategoryBrowser Infrastructure Headless browser platform for scraping, PDFs, and screenshots
CategoryBrowser Infrastructure Browser automation for agents, and a verified Cloudflare bot
CategoryBrowser Infrastructure Browser-as-a-service with natural-language control and fast launches
CategoryBrowser Infrastructure Natural-language cloud browser with SOC2 and HIPAA compliance
CategoryBrowser Infrastructure Hands any website to an AI agent, the top browser-automation repo
CategoryBrowser Infrastructure API for the only independent Western search index at scale
CategoryIndependent Web Indexes Nonprofit web archive of 9.5 petabytes behind most major LLMs
CategoryIndependent Web Indexes UK search engine with its own index and a no-tracking pledge
CategoryIndependent Web Indexes Structured feeds of news, forum, and dark-web content for enterprise
CategoryIndependent Web Indexes Programmatic access to Yandex index, dominant across Russia
CategoryIndependent Web Indexes Open-source search you re-rank yourself, built on its own index
CategoryIndependent Web Indexes Search engine for the text-heavy old web Google buries
CategoryIndependent Web Indexes French GDPR-native search running its own partial index
CategoryIndependent Web Indexes Open, decentralized index aiming to rival Google and Bing
CategoryIndependent Web Indexes Long-running open-source search engine with its own web index
CategoryIndependent Web Indexes Vision and NLP parsing behind a 10-billion-entity knowledge graph
CategoryAgentic Extraction Python library that scrapes sites from plain-English prompts
CategoryAgentic Extraction TypeScript SDK for driving browsers with natural-language steps
CategoryAgentic Extraction Describe the data you need and get a REST API for any site
CategoryAgentic Extraction One API that generates self-updating extraction workflows
CategoryAgentic Extraction Automates multi-step data extraction across many websites
CategoryAgentic Extraction Auto-generates extraction logic with no selectors or code
CategoryAgentic Extraction AI agent that navigates and extracts using vision and LLMs
CategoryAgentic Extraction Query language for pulling structured data from any web page
CategoryAgentic Extraction Turns unstructured web content into structured enterprise datasets
CategoryAgentic Extraction Point-and-click robots that extract and monitor any site
CategoryAgentic Extraction The original Python framework for large-scale web crawling
CategoryOpen Source Frameworks Apify's crawling library wrapping Playwright and Puppeteer
CategoryOpen Source Frameworks Open-source crawler shaped for RAG, the most-starred on GitHub
CategoryOpen Source Frameworks Fast concurrent scraping framework for Go with a callback API
CategoryOpen Source Frameworks Strips nav, ads, and chrome to leave the main article text
CategoryOpen Source Frameworks The library behind Firefox Reader Mode, extracting article text
CategoryOpen Source Frameworks Benchmark testing browser agents on 153 tasks across 144 live sites
812 long-horizon web tasks on reproducible self-hosted sites
643 live-web tasks measuring end-to-end multimodal agents
2,350 tasks across 137 sites measuring cross-domain transfer
369 computer-use tasks spanning Ubuntu, Windows, and macOS