crawlab-team/crawlab
Distributed web crawler admin platform for spiders management regardless of languages and frameworks. 分布式爬虫管理平台,支持任何语言和框架 observed · 2026-08-28
Health v2 · maintenance only
52/100
- Activity 66
- Release rhythm 8
- Longevity 100
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.
- gap_med: n/a
- age_days: 2761
- days_rel: n/a
- days_push: 205
- n_releases_24m: 0
Adoption not part of the score
12262 stars · 1888 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded
Crawlab is a Go-based distributed web crawler management platform with a web UI for managing, scheduling, and monitoring spiders written in any language or framework (Scrapy, Puppeteer, Selenium, etc.). It supports multi-node clusters over gRPC, a built-in Monaco code editor, task logs, and automatic storage of scraped data into MongoDB or SQL databases via native ORM.
Use cases
- manage and schedule web crawlers from a central dashboard
- run distributed scraping across multiple worker nodes
- monitor spider task logs and results in real time
- store scraped data into MySQL, PostgreSQL, or MongoDB automatically
- edit and deploy spider scripts in the browser
- replace scrapyd with a multi-language crawler platform
When to choose
- you run many spiders in different languages or frameworks and need unified management
- you need distributed crawling with master/worker nodes and scheduling
- you want a self-hosted admin UI with logs, results, and database integration
When to avoid
- you only need a single simple scraper script without orchestration
- you want a scraping framework/library to write crawler code rather than a management platform
- you cannot run Docker or MongoDB in your environment
Facets
application · maturity active
web-scraping scheduling monitoring logging self-hosted gui orm database crawlers web-development self-hosted self-hosted cross-platform go web-crawler spider-management distributed-crawling scrapy scrapyd task-scheduling grpc monaco-editor data-extraction data-engineering automation docker web-server linux
5 sources
- readme: https://github.com/crawlab-team/crawlab · fetched 2026-08-28 · 87f15e160d04
- homepage: https://www.crawlab.cn · fetched 2026-08-29 · c5b65147607f
- site_page: https://docs.crawlab.io · fetched 2026-08-29 · df0bdf97e265
- site_page: https://docs.crawlab.io/docs/whats-new · fetched 2026-08-29 · 68f37fba5888
- site_page: https://www.crawlab.io/ · fetched 2026-08-29 · c5b65147607f
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| crawlab-team/crawlab | main | 52 |
For agents
markdown · JSON · MCP: product_card(name="crawlab-team/crawlab")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem