Logo
Deploy Now

Stars

17,248

Forks

1,487

Watchers

87

Developer links

Maxun

With over 16,800 GitHub stars and growing rapidly, Maxun has become the go-to open-source platform for teams who need structured web data without writing scrapers. The TypeScript-based platform provides a no-code visual recorder that captures point-and-click interactions in real-time browser sync, automatically generating reusable extraction robots that handle pagination, infinite scrolling, and dynamic content. LLM-powered extraction accepts natural language prompts like "Extract 10 companies from the Y Combinator website" without requiring a URL — Maxun identifies the source and performs the extraction autonomously. The platform handles authentication-protected pages, adapts automatically to website layout changes through self-healing selectors, and exports directly to Google Sheets, Airtable, or any destination via webhooks. Robots run on configurable schedules with cron-based timing, turning any website into a perpetually fresh RESTful API endpoint. The crawl engine discovers and processes linked pages across entire domains with configurable depth and URL filtering, while the search capability runs automated queries across multiple engines. Official Node.js and Python SDKs provide programmatic control over robot creation, execution, and data retrieval, with MCP integration enabling direct connection to AI tools like Claude. The n8n community node enables workflow automation without custom code. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. AGPLv3 licensed.

Maxun
Maxun
Maxun
Maxun
Maxun

Benefits

  • Zero-Code Data Extraction
  • Visual point-and-click recorder captures browser interactions and generates reusable extraction robots without writing any code, making web scraping accessible to non-technical teams.
  • AI-Powered Smart Extraction
  • LLM integration accepts natural language descriptions of desired data and autonomously identifies sources, navigates pages, and returns structured results without manual configuration.
  • Self-Healing Website Adaption
  • Robots automatically adapt to website layout changes through intelligent selector recovery, eliminating maintenance overhead when target sites update their HTML structure.
  • Scheduled Automated Pipelines
  • Cron-based scheduling runs extraction robots at configured intervals with webhook notifications on completion, turning any website into a continuously updated data source.

Features

  • Visual Recorder
  • Real-time browser sync captures clicks, scrolls, form fills, and navigation actions to generate extraction robots that replay across pagination and dynamic content.
  • Crawl & Search Engine
  • Domain-wide crawling with depth control discovers linked pages automatically, while search mode runs queries across engines and extracts results as structured metadata.
  • API & SDK Access
  • Official Node.js and Python SDKs with full programmatic control over robot creation, scheduling, execution, and data retrieval through documented RESTful endpoints.
  • MCP Integration
  • Model Context Protocol support enables AI tools like Claude to trigger extractions, crawls, and searches directly through the Maxun server connection.
  • Export Integrations
  • Direct export to Google Sheets, Airtable, and custom destinations via webhooks with n8n community node for no-code workflow automation pipelines.