Webclaw
未认领Fast, local-first web content extraction for LLMs. Scrape, crawl, extract structured data — all from Rust. CLI, REST API, and MCP server.
aiai-agentsai-scrapingclicrawlerdata-extractionfirecrawl-alternativehtml-to-markdownllmmarkdownmcpmcp-serverragrustself-hostedtls-fingerprintingweb-crawlerweb-extractionweb-scraperweb-scraping
安装
$
npx create-webclaw工具(10 个)
batch
Scrape multiple URLs in parallel
brand
Extract colors, fonts, logos, and metadata
crawl
Follow same-origin links and extract discovered pages
diff
Compare page content snapshots
extract
Convert page content into structured data
map
Discover URLs without extracting every page
research
Multi-source research workflow
scrape
Extract one URL as markdown, text, JSON, LLM format, or HTML
search
Search the web and scrape results
summarize
Summarize a page