extract_structured_data
Extract structured data using CSS selectors or LLM
How to use it
extract_structured_data is exposed by the Crawl MCP MCP server. Add the server to your MCP client (Claude Desktop, Cursor, Windsurf and others), and the extract_structured_data tool becomes available to the model automatically. See the full listing for setup details and every tool this server provides.
Install Crawl MCP
uvx --from git+https://github.com/walksoda/crawl-mcp crawl-mcpOther tools in Crawl MCP (18)
Crawl multiple URLs with fallback (max 3 URLs)
Extract transcripts from multiple YouTube videos (max 3)
Perform multiple Google searches (max 3)
Extract web page content with JavaScript support
Crawl with fallback strategies for anti-bot sites
Crawl multiple pages from a site with configurable depth
Process large content with chunking and BM25 filtering
Extract entities (emails, phones, etc.) from web pages
Extract YouTube video comments with pagination
Extract YouTube transcripts with timestamps
Get available search genres
Get supported file formats and capabilities
Get YouTube video metadata and transcript availability
Extract specific data from web pages using LLM
Multi-URL crawl with pattern-based config (max 5 URL patterns)
Convert PDF, Word, Excel, PowerPoint, ZIP to markdown
Search Google and crawl top results
Search Google with genre filtering