# getmaxun/maxun

🔥 The open-source no-code platform for web scraping, crawling, search and AI data extraction • Turn websites into structured APIs in minutes 🔥

Repository: https://github.com/getmaxun/maxun
Canonical: https://ross.abutalabs.com/products/maxun
Homepage: https://www.maxun.dev
Language: TypeScript
License: AGPL-3.0
License Family: copyleft
Topics: automation, no-code, scraper, web-scraper, web-scraping, api, browser-automation, playwright, self-hosted, robotic-process-automation, rpa, agents, data-extraction, webscraping, nocode, crawler, crawling, web-search
Last push: 2026-08-26T22:43:51+00:00

## Health v2 (maintenance only)
Score: 94/100 (v2, computed 2026-09-03T02:20:16.233290+00:00)
- activity 99, release rhythm 99, longevity 74
- inputs: {"age_days": 1045, "days_push": 7, "days_rel": 8, "gap_med": 13.5, "n_releases_24m": 47}
- flags: none
- formula: round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10)

## Adoption (not part of the score)
Stars 17299, forks 1495 (observed 2026-08-28T04:11:17.930796+00:00)

## What it is
Maxun is an open-source no-code platform for web scraping, crawling, search, and AI-powered data extraction that turns websites into structured APIs. It offers a visual browser recorder, scheduled monitors, LLM-driven extraction agents, and self-hostable deployment with API, SDK, CLI, and MCP interfaces.

## Use cases
- scrape product prices from an e-commerce site without writing code
- turn a website into a structured REST API
- crawl an entire site and export pages as markdown for LLM training
- extract data from pages behind a login on a schedule
- monitor a webpage for changes hourly and get webhook alerts
- use natural language to extract structured data from complex sites
- convert PDFs into clean markdown or lists of links

## When to choose
- you want no-code visual scraping with scheduled runs and API/webhook delivery
- you need to handle JavaScript-heavy or authenticated sites via real browser automation
- you want a self-hostable open-source alternative to proprietary scrapers
- you need LLM-ready markdown extraction and AI agents for complex sites

## When to avoid
- you need a lightweight programmatic scraping library embedded in your own codebase
- you require high-volume scraping without credit-based limits and prefer writing custom scripts
- you need non-web data sources or heavy ETL transformations beyond extraction

## Facets
- artifact type: application
- maturity: active
- function: web-scraping, agent-framework, rag, workflow-automation, scheduling, api-framework, ocr
- domain: crawlers, artificial-intelligence, self-hosted
- platform: self-hosted, cli
- tags: no-code, rpa, browser-automation, playwright, data-extraction, structured-api, stealth-mode, mcp, web-crawler, automation, data-engineering, web-server, docker, nodejs

## Member repositories
- getmaxun/maxun (main) score 94

## Provenance
- Observed fields: from GitHub, fetched 2026-08-28T04:11:17.930796+00:00.
- Health v2: computed from the inputs above; adoption is never an input.
- Inferred fields (summary, facets, guidance): AI-extracted, prompt v1, taxonomy v1, on 2026-08-29T17:03:11.753774+00:00, confidence not recorded.
  - readme: https://github.com/getmaxun/maxun (fetched 2026-08-28T04:11:17.930796+00:00, sha d88881b8f7d7)
  - homepage: https://www.maxun.dev (fetched 2026-08-29T08:01:41.414503+00:00, sha 21c0c8d845a6)
  - site_page: https://docs.maxun.dev/ (fetched 2026-08-29T08:01:41.426108+00:00, sha 6634c36876db)
  - site_page: https://docs.maxun.dev (fetched 2026-08-29T08:01:41.429626+00:00, sha 6634c36876db)
  - site_page: https://www.maxun.dev/about (fetched 2026-08-29T08:01:41.431744+00:00, sha e8f1330163e4)
  - site_page: https://www.maxun.dev/pricing (fetched 2026-08-29T08:01:41.427950+00:00, sha 7d817d183f36)
- Data as of 2026-08-30T08:39:29.467469+00:00.
