Welcome
Turn any website into live, clean structured data with Maxun
Maxun is the open-source web data platform for turning the web into structured, reliable data.
It supports extraction, crawling, scraping, monitoring and search - designed to scale from simple use cases to complex, automated workflows.
What Maxun Enables
| Extraction | Turn websites into structured datasets and APIs, even across dynamic pages and authenticated flows. Learn More. |
| Crawling | Automatically discover and collect data across entire websites with intelligent link following and scoped control. Learn More. |
| Scraping | Convert full webpages into clean Markdown, HTML and capture screenshots. Learn More. |
| Search | Run programmatic web searches and extract results as metadata or full content, with time-based filtering. Learn More. |
| Monitoring | Track websites over time, detect changes, and get notified when they occur. Learn More. |
| Document Extraction & Parsing | Extract and parse structured data from documents (PDF, CSV, XLSX, DOCX, JPG and PNG). Learn More. |
How To Use Maxun
Maxun is no-code by default, with APIs, CLI, MCP, and SDKs for deeper integrations without changing how extractions are defined.
| Dashboard | Build, run, and manage robots through Maxun's visual interface. No coding required. |
| API | Integrate Maxun's scraping and extraction capabilities into your applications. Learn More. |
| SDK | Use Maxun programmatically for scraping, extraction, automation and more. Learn more. |
| CLI | Create robots, trigger runs, and retrieve data from your terminal. Learn more. |
| MCP | Connect Maxun to AI agents through the Model Context Protocol. Learn more. |
Sponsors
Continue to the next chapter to understand how Maxun works and how these capabilities are powered.