jaypyles/Scraperr
Self-hosted webscraper. observed · 2026-08-28
Health v2 · maintenance only
10/100
- Activity 46
- Release rhythm 40
- Longevity 56
Flags: archived
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.
- gap_med: 2
- age_days: 788
- days_rel: 416
- days_push: 325
- n_releases_24m: 20
Adoption not part of the score
4910 stars · 249 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded
Scraperr is a self-hosted web scraping application with a web UI that lets users scrape websites without writing code, using XPath-based extraction, job queues, and domain spidering. It is built with FastAPI, Playwright, and Next.js, and deploys via Docker or Helm.
Use cases
- scrape websites without writing code
- extract data from web pages using xpath
- crawl all pages within a domain
- download images and videos from scraped pages
- export scraped data to csv or markdown
- manage a queue of scraping jobs
- self-host a web scraper with docker
When to choose
- you want a no-code, self-hosted scraping tool with a UI
- you need XPath-based extraction and domain-wide spidering
- you want job queues, notifications, and CSV/markdown export out of the box
- you deploy with Docker or Kubernetes/Helm
When to avoid
- you need a programmatic scraping library to embed in your own code
- you require large-scale distributed crawling beyond a single self-hosted instance
- you need scraping of sites requiring complex authentication or heavy JavaScript interaction beyond Playwright defaults
Facets
application · maturity active
web-scraping self-hosted gui data-visualization scheduling web-development crawlers self-hosted self-hosted python xpath playwright fastapi nextjs no-code-scraping helm data-export automation data-engineering docker kubernetes web-server
2 sources
- readme: https://github.com/jaypyles/Scraperr · fetched 2026-08-28 · 390f513825a1
- homepage: https://scraperr-docs.pages.dev/ · fetched 2026-08-29 · a78105121c6a
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| jaypyles/Scraperr | main | 10 |
For agents
markdown · JSON · MCP: product_card(name="jaypyles/Scraperr")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem