Firecrawl
io.mcp-cn.firecrawl-mcp
基于Firecrawl的智能网页爬取和内容提取MCP工具
“crawl scrape firecrawl fetch” 共 610 个结果
io.mcp-cn.firecrawl-mcp
基于Firecrawl的智能网页爬取和内容提取MCP工具
v0.1.2
io.github.pipeworx-io/firecrawl
Firecrawl MCP,封装 Firecrawl API(firecrawl.dev)。
v1.0.3
com.mcparmory/firecrawl
大规模抓取、爬取网页并提取结构化数据
io.smithery.intake-triage.steadyfetch
可靠的网页抓取 MCP 服务器,内置重试逻辑、熔断器模式、缓存与反爬绕过。可按原始 HTML 或便于 LLM 消费的干净 Markdown 抓取 URL。还提供域名健康检查与缓存管理工具。
io.smithery.reyd8777.framefetch
# FrameFetch **一次 API/MCP 调用,智能体即可获得跨 6 个平台的干净视频数据**——元数据与洞察、Whisper 转写文本,以及参数化抽帧(自选 fps 或精确时间戳 → 推送到 S3)。支持 YouTube(含 Shorts)、TikTok、Reddit、Instagram、Pinterest。 智能体优先设计:类型化错误、失败退款、结果缓存。通过 x402(Base 上的 USDC)或 Stripe 按次付费。 ## 接口 - POST /v1/extract——一次调用组合元数据/洞察/转写/抽帧中任意几项 - POST /v1/metadata · /v1/transcript · /v1/frames——快捷接口 - GET /v1/platforms——能力矩阵 · POST /v1/keys——免费 key 加额度 ## 示例 curl -X POST https://framefetch.net/v1/extract -H “Authorization: Bearer <key>” -H “Content-Type: application/json” -d {“url”:“https://youtu.be/...”,“fields”:[“metadata”,“transcript”]} https://framefetch.net
io.smithery.axel-belfort.web-scraper
Web content extraction API for AI agents. Scrape any URL and get clean, structured Markdown content with navigation, ads, and scripts stripped. Full JavaScript rendering via headless Chromium. Single and batch (10 URLs) modes. Built for RAG pipelines and AI research. Tools: web_scrape_to_markdown (single), web_scrape_batch (up to 10 URLs). Use this for RAG ingestion, research, content analysis, data extraction, or competitive intelligence. IMPORTANT: For screenshots/PDFs of pages, use capture_screenshot instead. For SEO analysis, use seo_audit_page. Returns: {markdown, title, wordCount, links[]}. No API key required — x402 micropayment $0.005/call on Base L2.
io.smithery.axel-belfort.twitter-scraper
面向 AI 智能体的 Twitter/X 抓取 API。抓取公开资料(简介、数据、认证)、用户推文(文本、互动、媒体)与搜索结果——全部不需要 Twitter API key。输出结构化的 JSON,可直接分析。 工具:twitter_scrape_profile、twitter_search_tweets、twitter_get_user_tweets。 适合社媒监控、网红调研、情感分析、竞品情报,或搭建社交看板。是为 AI 智能体准备的社交情报层。重要:若要其他社交网络,请使用 social_lookup_profile。 返回含资料、推文与互动指标的结构化 JSON。无需 API key——在 Base L2 上以 x402 微支付,每次 0.005 美元。
io.smithery.hshintelligence.agentscrape
**面向 AI 智能体的按次付费网页抓取——无需注册、无需 API key,只用 USDC。** AgentScrape 是可用于生产的 MCP 服务器,为智能体提供 6 个付费工具,用于网页抓取、结构化抽取、截图与元数据——全部运行在 Cloudflare Workers,全部通过 x402 在 Base 上以 USDC 结算。 ## 工具 - **`scrape_webpage`**——按 markdown、html、text 或 json 抓取 - **`extract_structured_data`**——用自然语言提示加 JSON schema 做 AI 抽取 - **`screenshot_webpage`**——PNG 截图,视口可配置 - **`extract_metadata`**——OG、Twitter Card、JSON-LD - **`create_browser_session`**——为多步流程创建持久会话 - **`run_workflow`**——多步原子执行 ## 定价 Base 主网上每次调用 0.001–0.008 USDC。无订阅,无 API key,通过 x402 按请求付费。 **免费额度:**每个钱包每 30 天 10 次调用,无需注册。 ## 技术栈 Cloudflare Workers · Hono · `@x402/hono` v2 · MCP Streamable HTTP · xpay.sh facilitator · Browser Rendering MIT 许可。
io.smithery.scrapegraphai-inc.sgai
ScrapeGraphAI MCP 服务器是一个可用于生产的 Model Context Protocol(MCP)服务器,把大语言模型(LLM)接到 ScrapeGraph AI API 上。让 Claude、Cursor 这类 AI 助手直接通过自然语言交互完成 AI 驱动的网页抓取、调研与爬取。
v0.7.2
io.github.fetchsandbox/mcp
面向 agent 的确定性验证引擎:为修复提供证明——旧代码上失败、新代码上通过。
v0.1.3
io.github.imfurkana/unrendered-ai-crawler-checker
AI crawler checker: shows which page content GPTBot and ClaudeBot can't read without JavaScript.
v1.0.0
com.meridianlabssoftware/eu-uk-tenders-scraper
Public tenders and contract awards from EU TED and UK Find a Tender, with buyer, value and deadline.
v6.19.2
io.github.mysleekdesigns/crawlforge-mcp-server
网页抓取、遍历、深度调研与自主信息抽取,31 个 MCP 工具,输出干净的 Markdown/JSON。
v2.0.0
dev.scrapewhale/scrapewhale
Web data for marketing agents: traffic, ads, social profiles, search, and any URL as markdown.
v1.9.0
io.github.xberg-io/crawlberg
通过本地 CLI 抓取、爬取并映射网站,输出 Markdown 或 JSON。
v2.3.1
io.github.nolindnaidoo/scrape-le
Analyse robots.txt content and report whether a path may be crawled.
v1.0.1791044090
com.saastemly/crawl4agent
crawl4ai-compatible web crawler API: POST a URL.
v1.0.0
io.github.saulius876-lgtm/telegram-channels
Public Telegram channel posts without login: text, views, reactions, media, exact subscribers.
v1.0.0
io.github.saulius876-lgtm/trustpilot-reviews
Trustpilot reviews for any company domain: ratings, text, replies, and unanswered low-star reviews.
v1.0.0
io.github.saulius876-lgtm/reddit-search
Search Reddit posts by keyword and subreddit; find recommendation requests and brand mentions.
v1.0.0
io.github.saulius876-lgtm/naukri-jobs
Search Naukri.com jobs in India: titles, companies, salaries, skills, and which companies hire most.
v1.0.1
io.github.tidytools/tidytools-apify-tools
Enrich Google Maps leads, crawl docs to RAG Markdown, transcribe podcasts, check AI crawler access.
v3.27.3
io.github.firecrawl/firecrawl-mcp-server
Firecrawl 的 MCP server——网页搜索、网页抓取,以及生物医学/arXiv 论文检索。
v1.23.0
io.github.maximilianfeix/proxy-scraper
Free proxies that actually work: verified HTTP/SOCKS proxies, and pages fetched through them