Turn any website into clean markdown or structured JSON — crawling, JavaScript rendering, and extraction in one API.
Scraping
Extracting data from websites at scale, handling rendering, pagination and blocks.
10 apps, 2 skills and 8 MCP servers tagged Scraping.
Apps
Web data extraction APIs for AI agents — 75+ ready-made scrapers for LinkedIn, Amazon, Google Maps, Reddit and X, priced per result.
One API to search, scrape, crawl, map and monitor the web, returning clean structured text that AI agents and RAG pipelines can use directly.
Describe the data you want in plain English and BrowserAct builds a reusable scraper that runs in a real browser, with proxies and CAPTCHA handled.
Cloud browser automation for AI agents: describe a workflow in plain English and Airtop compiles it into a deterministic agent that logs in, browses and acts at scale.
One REST API that scrapes, crawls, and extracts the web into clean Markdown and schema-validated JSON for AI agents.
A marketplace of thousands of ready-made scrapers and automation tools, plus the cloud to run your own.
Headless browser infrastructure for AI agents — sessions, stealth, proxies, and live view, without running Chrome yourself.
The open-source library that lets AI agents actually use websites — clicking, typing, and reading like a person.
A spreadsheet that enriches itself — pull data from 100+ sources, research with AI, and run go-to-market plays.
Skills
Skill: Browserbase Browser Automation
by Browserbase
Browserbase's official skill for driving a real browser from an agent via the browse CLI - navigate, snapshot, fill forms and click, locally or in a cloud session.
Skill: Firecrawl Build: Scrape
by Firecrawl
Integrate Firecrawl's /scrape endpoint into product code — turn a known URL into markdown, HTML, links, metadata or screenshots, with sensible defaults baked in.
MCP servers
Apify's official MCP server — turn thousands of ready-made scrapers into tools your assistant can call.
Official Crawlbase MCP server giving agents live web access — fetch any URL as raw HTML, clean Markdown or a screenshot, with JS rendering and proxy rotation handled for you.
MCP: wigolo
by KnockOutEZ
Local-first web MCP server for AI agents — search, fetch, crawl, extract and deep research with no API keys, no cloud round-trip and no per-query bill.
MCP: Bright Data
by Bright Data
Bright Data's official MCP server — search, scrape, and extract web data through a large proxy network.
MCP: Puppeteer
by Model Context Protocol
Browser automation for AI clients — navigate, click, type, screenshot, and run JavaScript in a real page.
MCP: Exa Search
by Exa
Real-time web search and crawling built for AI — neural search with configurable tools and live content retrieval.
MCP: Firecrawl
by Firecrawl
Search, scrape, crawl, and extract structured data from the web — with JS rendering, batch scraping, and LLM-powered extraction.
MCP: Fetch
by Model Context Protocol
Fetch a URL and convert its contents to clean Markdown for efficient LLM consumption.
Related tags
Tags that appear alongside this one, ranked by how often.