webclaw
Extract websites into clean Markdown, JSON, or LLM-ready context through CLI, MCP, or API.
- Stars
- 1,729
- Forks
- 181
- Updated
- Updated Jul 13, 2026
Extract websites into clean Markdown, JSON, or LLM-ready context through CLI, MCP, or API.
webclaw provides a local-first way to scrape, crawl, map, and extract websites for agents and RAG systems. Its core CLI and MCP workflow can run without an account for many pages, while a hosted API is available for bot-protected sites, JavaScript rendering, asynchronous crawl jobs, search, and production monitoring.
Resource types
Use cases
Platforms
Manage and query Google NotebookLM through a CLI or MCP server.
Fetch structured data from dozens of websites with one CLI command.
Runtime
Protocols & integrations
Capabilities
Audience
Public GitHub facts last synced Jul 14, 2026.
Search, scrape, and browse the web from a terminal or coding agent.
Turn an expert’s public work into reusable SKILL.md or AGENTS.md guidance.