Ross ROSS = Recommend OSS · open-source software intelligence for agents

crawlab-team/crawlab

Distributed web crawler admin platform for spiders management regardless of languages and frameworks. 分布式爬虫管理平台,支持任何语言和框架 observed · 2026-08-28

github.com/crawlab-team/crawlab · homepage · Go · BSD-3-Clause (permissive) observed · 2026-08-28

Health v2 · maintenance only

52/100

  • Activity 66
  • Release rhythm 8
  • Longevity 100
How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 2761
  • days_rel: n/a
  • days_push: 205
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

12262 stars · 1888 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded

Crawlab is a Go-based distributed web crawler management platform with a web UI for managing, scheduling, and monitoring spiders written in any language or framework (Scrapy, Puppeteer, Selenium, etc.). It supports multi-node clusters over gRPC, a built-in Monaco code editor, task logs, and automatic storage of scraped data into MongoDB or SQL databases via native ORM.

Use cases

  • manage and schedule web crawlers from a central dashboard
  • run distributed scraping across multiple worker nodes
  • monitor spider task logs and results in real time
  • store scraped data into MySQL, PostgreSQL, or MongoDB automatically
  • edit and deploy spider scripts in the browser
  • replace scrapyd with a multi-language crawler platform

When to choose

  • you run many spiders in different languages or frameworks and need unified management
  • you need distributed crawling with master/worker nodes and scheduling
  • you want a self-hosted admin UI with logs, results, and database integration

When to avoid

  • you only need a single simple scraper script without orchestration
  • you want a scraping framework/library to write crawler code rather than a management platform
  • you cannot run Docker or MongoDB in your environment

Facets

application · maturity active

web-scraping scheduling monitoring logging self-hosted gui orm database crawlers web-development self-hosted self-hosted cross-platform go web-crawler spider-management distributed-crawling scrapy scrapyd task-scheduling grpc monaco-editor data-extraction data-engineering automation docker web-server linux

5 sources

Member repositories

RepositoryRoleHealth v2
crawlab-team/crawlabmain52

For agents

markdown · JSON · MCP: product_card(name="crawlab-team/crawlab")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem