Home / Topics / Web Scraping MCP servers

Topic · updated 2026-10-05

Web Scraping MCP servers

164 Web Scraping MCP servers indexed on awesomemcp.in. Every entry links to a live GitHub repository with a known author, is de-duplicated across the official MCP registry and community lists, and carries an automated vulnerability-check badge. Built by @apify, @firecrawl, @fastcrw, @Xquik-dev, @Sriram-PR and 145 other authors.

164servers
158scan passed
23official
93one-click install

Top 60 Web Scraping MCP servers

  1. by @apify · official · TypeScript · apify/apify-mcp-server

    Extract data from any website with thousands of scrapers, crawlers, and automations on Apify Store ⚡

    Scan passed
    ★ 9,728
    2026-10-05
  2. by @firecrawl · official · TypeScript · firecrawl/firecrawl-mcp-server

    MCP server for Firecrawl — web search, scraping, and biomedical/arXiv paper search.

    Scan passed
    ★ 7,554
    2026-10-05
  3. by @fastcrw · official · Rust · fastcrw/crw

    Open-source web scraper for AI agents with scrape, crawl, and map tools

    Scan passed
    ★ 1,102
    2026-10-04
  4. X Twitter scraper & Twitter API alternative. Search, monitor, publish & manage X accounts.

    Scan passed
    ★ 210
    2026-10-05
  5. Crawl documentation sites into local corpora agents can search, read, and diff fully offline.

    Scan passed
    ★ 102
    2026-10-05
  6. MCP server for AI agents to search major Chinese internet platforms with local-first login support.

    Scan passed
    ★ 53
    2026-07-29
  7. Free proxies that actually work: verified HTTP/SOCKS proxies, and pages fetched through them

    Scan passed
    ★ 40
    2026-10-05
  8. Live HKEx (Hong Kong Stock Exchange) regulatory filings for AI agents.

    Scan passed
    ★ 16
    2026-10-05
  9. Fetch any public web page through managed proxies, with optional JS rendering and extraction rules.

    Scan passed
    ★ 11
    2026-10-05
  10. US Congress stock trades and financial disclosures by member, ticker, or date, hosted MCP.

    Scan passed
    ★ 9
    2026-09-24
  11. Web scraping: clean content extraction, CSS-selector scraping, Playwright screenshots, link and metadata extraction, and Google search.

    Scan passed
    ★ 8
    2026-09-11
  12. Web scraping MCP — extract clean markdown, links, and metadata from any URL.

    Scan passed
    ★ 6
    2026-08-10
  13. Google Hotels prices, ratings, reviews, and photos via an Apify Actor, hosted MCP.

    Scan passed
    ★ 3
    2026-09-24
  14. Scrape Amazon products, search, and async batch ASIN lookups across 20 marketplaces

    Scan passed
    ★ 2
    2026-06-02
  15. Pay-per-call web scraping for AI agents via x402 on Base USDC. Six tools, no signup.

    Scan passed
    ★ 2
    2026-06-22
  16. Twitter/X scraping for AI agents: profiles, tweets, followers, trends, lists and communities.

    Scan passed
    ★ 2
    2026-09-30
  17. Search bounded public Bluesky keywords, handles, mentions, and hashtags.

    Scan passed
    ★ 2
    2026-09-16
  18. Three MCP tools: fetch_page, fetch_pages_batch, search_web. Ad-free Markdown for AI agents.

    Scan passed
    ★ 2
    2026-08-26
  19. Detects GTM hiring activity from company career pages via Greenhouse, Lever, Ashby. Clay-ready.

    Scan passed
    ★ 1
    2026-10-05
  20. Detects a company CRM, sequencer, and marketing automation from its public website. Clay-ready.

    Scan passed
    ★ 1
    2026-10-05
  21. Apify MCP — run web-scraping Actors and fetch their dataset results.

    Scan passed
    ★ 1
    2026-09-25
  22. Firecrawl MCP — wraps the Firecrawl API (firecrawl.dev) for web

    Scan passed
    ★ 1
    2026-09-26
  23. Web scraping with escalating fetch tiers on detected blocks, returning clean Markdown or saving large scrapes to SQLite for read-only SQL queries.

    Scan passed
    ★ 1
    2026-08-31
  24. Audit any site for 50+ AI crawlers, generate llms.txt, robots.txt and schema, track AI mentions.

    Scan passed
    ★ 0
    2026-09-30
  25. 40 Apify public-data tools for leads, news, SEO, jobs, SEC, procurement, and registries.

    Scan passed
    ★ 0
    2026-09-18
  26. Web scraping, Llama 3 digests and Twitter intelligence, paid per call in USDC via x402 on Base.

    Scan passed
    ★ 0
    2026-09-16
  27. by @devflowinc · official · TypeScript · devflowinc/trieve

    Crawl, embed, chunk, search and retrieve information from datasets through Trieve.

    Scan passed
    ★ 2,718
    2026-01-25
  28. by @brightdata · official · TypeScript · brightdata/brightdata-mcp

    Bright Data's Web MCP server enabling AI agents to search, extract & navigate the web

    Scan passed
    ★ 2,660
    2026-09-17
  29. by @nottelabs · official · Python · nottelabs/notte

    Leverage Notte Web AI agents & cloud browser sessions for scalable browser automation & scraping workflows

    Needs review
    ★ 2,014
    2026-10-05
  30. by @graphlit · official · TypeScript · graphlit/graphlit-mcp-server

    Ingest content from Slack, Discord, websites, Google Drive, Linear or GitHub into a Graphlit project, then search and retrieve relevant knowledge.

    Scan passed
    ★ 378
    2026-01-12
  31. by @superagents-lab · official · TypeScript · superagents-lab/search1api-mcp

    Web search, news, page retrieval, sitemaps, and trending topics through Search1API.

    Scan passed
    ★ 173
    2026-09-29
  32. Web search, browser automation, scraping, crawling and CAPTCHA solving for AI agents.

    Scan passed
    ★ 169
    2026-09-08
  33. by @supadata-ai · official · TypeScript · supadata-ai/mcp

    Official MCP server for Supadata - YouTube, TikTok, X and Web data for makers.

    Scan passed
    ★ 63
    2026-10-05
  34. by @crawlbase · official · JavaScript · crawlbase/crawlbase-mcp

    Enables AI agents to access real-time web data with HTML, markdown, and screenshot support. SDKs: Node.js, Python, Java, PHP, .NET.

    Needs review
    ★ 58
    2026-04-23
  35. Web scraping tools with Chromium JS rendering, rotating proxies, and AI question answering.

    Needs review
    ★ 44
    2026-09-27
  36. by @Decodo · official · TypeScript · Decodo/mcp-server

    Enable your AI agents to scrape and parse web content dynamically, including geo-restricted sites

    Needs review
    ★ 37
    2026-10-05
  37. by @serkan-ozal · official · TypeScript · serkan-ozal/driflyte-mcp-server

    Discover available topics and explore up-to-date, topic-tagged web content. Search to surface the…

    Needs review
    ★ 11
    2025-10-02
  38. by @browserless · official · TypeScript · browserless/browserless-mcp

    Headless browser automation and web scraping via the Browserless smart scraper API.

    Scan passed
    ★ 7
    2026-10-05
  39. by @any4ai · official · TypeScript · any4ai/anycrawl-mcp-server

    AnyCrawl MCP Server, Powerful web scraping and crawling for Cursor, Claude, and other LLM clients via the Model Context Protocol (MCP).

    Needs review
    ★ 6
    2026-09-10
  40. by @usestring · official · TypeScript · usestring/string-ai-mcp

    The most accurate web access API. Stop getting blocked.

    Scan passed
    ★ 2
    2026-10-05
  41. by @getanyapi-com · official · TypeScript · getanyapi-com/mcp

    Hundreds of scraping & data APIs through one key. USD pay-per-request, normalized schemas, failover.

    Scan passed
    ★ 1
    2026-09-06
  42. Ten government-record evidence tools for property, healthcare, facilities, Norway, and routing.

    Scan passed
    ★ 0
    2026-10-01
  43. Live web access for agents: scrape, SERP search, crawl/map, 100+ collectors, datasets, proxies.

    Scan passed
    ★ 0
    2026-10-03
  44. by @ToolTrace-io · official · TypeScript · ToolTrace-io/mcp-server

    Web tools for AI agents: scrape pages to Markdown, audit SEO, detect tech stacks, check sitemaps

    Scan passed
    ★ 0
    2026-09-08
  45. by @Citlyze · official · JavaScript · Citlyze/citlyze-mcp

    Track brand visibility in AI search with Citlyze: scores, tracked prompts, citations, competitor comparison, recommendations and AI crawler analytics.

    Scan passed
    ★ 0
    2026-08-19
  46. by @quantumproxies · official · TypeScript · quantumproxies/quantumproxies-mcp

    Residential-proxy web access: scrape pages to Markdown, SERP search, site mapping and crawling, ready-made collectors, and proxy endpoints from your plans.

    Scan passed
    ★ 0
    2026-10-01
  47. by @D4Vinci · Python · D4Vinci/Scrapling

    Web scraping with stealth HTTP, real browsers, and Cloudflare bypass. CSS selectors supported.

    Scan passed
    ★ 85,824
    2026-10-04
  48. AI-powered web scraping library that creates scraping pipelines using natural language.- ScrapeGraphAI

    Scan passed
    ★ 31,543
    2026-09-25
  49. Turn docs, GitHub repos, PDFs, videos, Confluence, Notion and Slack/Discord into AI skills and RAG knowledge, exportable to vector databases like Qdrant.

    Scan passed
    ★ 15,100
    2026-09-30
  50. Stealthy browser automation, testing, and web-scraping via CDP Mode.

    Scan passed
    ★ 13,048
    2026-10-02
  51. Dark web OSINT over Tor: search onion engines, scrape pages, report with your own model.

    Scan passed
    ★ 7,424
    2026-10-03
  52. Web scraping, crawling, extraction, summarization, diffing and research, using TLS fingerprinting to bypass anti-bot checks without a browser.

    Scan passed
    ★ 2,366
    2026-09-23
  53. Stealth browser for AI agents: fetch pages behind Cloudflare/DataDome/CAPTCHA, extract clean data.

    Scan passed
    ★ 725
    2026-10-05
  54. X/Twitter automation MCP: scrape, post, schedule, analyze, engage. No API key required.

    Scan passed
    ★ 569
    2026-10-05
  55. MCP server for integrating Omnisearch with LLMs

    Scan passed
    ★ 352
    2026-10-05
  56. Public Spotify metadata, lyrics and podcasts for LLM agents. No API key, read-only.

    Scan passed
    ★ 316
    2026-09-18
  57. Web-scraping toolkit with 22 tools for structured web data as JSON for AI agents.

    Scan passed
    ★ 261
    2026-09-25
  58. Search jobs and companies on Job Seek (jseek.co)

    Scan passed
    ★ 201
    2026-10-05
  59. Scrape, crawl, and map websites to Markdown or JSON via local CLI.

    Scan passed
    ★ 182
    2026-10-05
  60. Renders web pages into structured, agent-readable representations using headless Chromium.

    Scan passed
    ★ 180
    2026-09-11

Search all 164 in the interactive index →

Quick install

Add the most popular option to Claude Code:

claude mcp add --transport http apify-mcp-server https://mcp.apify.com/

FAQ

What is the best web scraping mcp server?

By GitHub stars, the most popular is apify-mcp-server by @apify (9,728 stars). awesomemcp.in lists 164 MCP servers for this topic, ranked by official status and popularity.

Are these MCP servers safe to install?

158 of 164 passed our automated vulnerability check (OSV.dev advisories for the published package, OpenSSF Scorecard, license, maintenance and provenance). The check is automated and does not guarantee 100% safety — review the code and permissions before installing.

How do I install apify-mcp-server?

With Claude Code: claude mcp add --transport http apify-mcp-server https://mcp.apify.com/. Every listing on awesomemcp.in includes an install command and, where available, an mcp.json snippet for Claude Desktop, Cursor and other MCP clients.

Can AI agents read this list?

Yes. The full catalog is available as JSON at https://awesomemcp.in/api/mcps.json and https://awesomemcp.in/api/skills.json, and as plain text at https://awesomemcp.in/llms.txt.

Related topics