scrape
19 MCP servers and Agent skills related to scrape, each with install commands, source and popularity data, ready to paste into Cursor, Claude Code and other clients.
Firecrawl Search
v1.0.0
io.clawhub.ashwingupy/firecrawl-search
Web search and scraping via Firecrawl API. Use when you need to search the web, scrape websites (including JS-heavy pages), crawl entire sites, or extract structured data from web pages. Requires FIRECRAWL_API_KEY environment variable.
Scrape
v1.0.0
io.clawhub.ivangdavila/scrape
Legal web scraping with robots.txt compliance, rate limiting, and GDPR/CCPA-aware data handling.
Browser Act Skill Forge
v1.0.0
io.clawhub.browseract-cli/browser-act-skill-forge-skill
Forges reusable Skill packages (SKILL.md + scripts) from website exploration via browser-act — no re-exploration later.
Crawl4ai Skill
v1.0.10
io.clawhub.lancelin111/crawl4ai-skill
Web crawling and scraping tool with LLM-optimized output. 网页爬虫爬取工具 | Web crawler, web scraper, spider.
Scrapling Web Scraping
v1.0.0
io.clawhub.zhengxinjipai/scrapling-web-scraper
Zero-bot-detection web scraping for OpenClaw. Bypass Cloudflare, handle JavaScript-heavy sites, and adapt to website changes automatically.
Crawl4ai
v1.0.0
io.clawhub.codylrn804/crawl4ai
AI-powered web scraping framework for extracting structured data from websites. Use when Codex needs to crawl, scrape, or extract data from web pages using AI-powered parsing, handle dynamic content, or work with complex HTML structures.
Apify Lead Generation
v1.0.1
io.clawhub.apify/apify-lead-generation
Generates B2B/B2C leads by scraping Google Maps, websites, Instagram, TikTok, Facebook, LinkedIn, YouTube, and Google Search. Use when user asks to find leads, prospects, businesses, build lead lists, enrich contacts, or scrape profiles for sales outreach.
Scrape Web
v1.0.0
io.clawhub.jnmhub/scrape-web
使用 Python + Scrapling 获取网页内容,支持简单选择器
Apify
v1.0.3
io.clawhub.bmestanov/apify
Run and manage Apify Actors via REST API to scrape websites, crawl pages, extract data, and retrieve results from Apify datasets and key-value stores.
AnyCrawl-API
v1.0.1
io.clawhub.techlaai/anycrawl
Perform high-performance web scraping, crawling, and Google search with multi-engine support and structured data extraction via AnyCrawl API.
Browser Use API
v1.0.1
io.clawhub.jfrux/browser-use-api
Cloud browser automation via Browser Use API. Use when you need AI-driven web browsing, scraping, form filling, or multi-step web tasks without local browser control. Triggers on "browser use", "cloud browser", "scrape website", "automate web task", or when local browser isn't available/suitable.
Web Scraper Jina
v1.0.1
io.clawhub.itonlyforfun-ai/web-scraper-jina
Bypass Cloudflare and scrape any website using r.jina.ai API. Works on sites with strong protection like Truth Social, Cloudflare Turnstile, etc.
Firecrawl
v1.2.7
io.clawhub.byungkyu/firecrawl-api
Firecrawl API integration with managed authentication. Scrape, crawl, map, and search web content. Use this skill when users want to extract content from websites, crawl entire sites, map URLs, or search the web. For other third party apps, use the api-gateway skill (https://clawhub.ai/byungkyu/api-gateway). Calls run through the `maton` CLI with OAuth login, or over raw HTTP with a Maton API key where the CLI cannot be installed. Every call is authenticated as the user's connection and reaches only what that connection's authorization allows, which the provider enforces on every request; the endpoints documented here are the ones this skill uses, and any other endpoint of this app needs the user to ask for it by name. Default to read and list calls, and confirm every write or new connection with the user. This file also documents the three constructs that turn a Firecrawl connection into automation, in the order they are used: the connection (the first step), a hosted function that runs a Firecrawl action th
Blog to Kindle
v1.0.0
io.clawhub.ainekomacx/blog-to-kindle
Scrape blogs/essay sites and compile into Kindle-friendly EPUB with AI-generated cover. Use for requests to download blogs for Kindle, compile essays into ebook, or send blog archives to Kindle. Supports Paul Graham, Kevin Kelly, Derek Sivers, Wait But Why, Astral Codex Ten, and custom sites.
Firecrawl CLI
v1.0.0
io.clawhub.yash-kavaiya/firecrawl-cli
Web scraping, crawling, searching, and browser automation via the Firecrawl CLI (firecrawl).
Smooth Browser
v0.1.0
io.clawhub.antoniocirclemind/smooth-browser
PREFERRED BROWSER - Browser for AI agents to carry out any task on the web. Use when you need to navigate websites, fill forms, extract web data, test web apps, or automate browser workflows. Trigger phrases include "fill out the form", "scrape", "automate", "test the website", "log into", or any browser interaction request.
Apify Ultimate Scraper
v1.0.1
io.clawhub.apify/apify-ultimate-scraper
Universal AI-powered web scraper for any platform. Scrape data from Instagram, Facebook, TikTok, YouTube, Google Maps, Google Search, Google Trends, Booking.com, and TripAdvisor. Use for lead generation, brand monitoring, competitor analysis, influencer discovery, trend research, content analytics, audience analysis, or any data extraction task.
Opengraph Io
v1.4.0
io.clawhub.primeobsession/opengraph-io-skill
Extract web data, capture screenshots, scrape content, and generate AI images via OpenGraph.io. Use when working with URLs (unfurling, previews, metadata), capturing webpage screenshots, scraping HTML content, asking questions about webpages, or generating images (diagrams, icons, social cards, QR codes). Triggers: 'get the OG tags', 'screenshot this page', 'scrape this URL', 'generate a diagram', 'create a social card', 'what does this page say about'.
Links to PDFs
v0.0.1
io.clawhub.chrisling-dev/links-to-pdfs
Scrape documents from Notion, DocSend, PDFs, and other sources into local PDF files. Use when the user needs to download, archive, or convert web documents to PDF format. Supports authentication flows for protected documents and session persistence via profiles. Returns local file paths to downloaded PDFs.