Ross ROSS = Recommend OSS · open-source software intelligence for agents

jaypyles/Scraperr

Self-hosted webscraper. observed · 2026-08-28

github.com/jaypyles/Scraperr · homepage · TypeScript · MIT (permissive) · archived observed · 2026-08-28

Health v2 · maintenance only

10/100

  • Activity 46
  • Release rhythm 40
  • Longevity 56

Flags: archived

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.

  • gap_med: 2
  • age_days: 788
  • days_rel: 416
  • days_push: 325
  • n_releases_24m: 20

Full methodology

Adoption not part of the score

4910 stars · 249 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded

Scraperr is a self-hosted web scraping application with a web UI that lets users scrape websites without writing code, using XPath-based extraction, job queues, and domain spidering. It is built with FastAPI, Playwright, and Next.js, and deploys via Docker or Helm.

Use cases

  • scrape websites without writing code
  • extract data from web pages using xpath
  • crawl all pages within a domain
  • download images and videos from scraped pages
  • export scraped data to csv or markdown
  • manage a queue of scraping jobs
  • self-host a web scraper with docker

When to choose

  • you want a no-code, self-hosted scraping tool with a UI
  • you need XPath-based extraction and domain-wide spidering
  • you want job queues, notifications, and CSV/markdown export out of the box
  • you deploy with Docker or Kubernetes/Helm

When to avoid

  • you need a programmatic scraping library to embed in your own code
  • you require large-scale distributed crawling beyond a single self-hosted instance
  • you need scraping of sites requiring complex authentication or heavy JavaScript interaction beyond Playwright defaults

Facets

application · maturity active

web-scraping self-hosted gui data-visualization scheduling web-development crawlers self-hosted self-hosted python xpath playwright fastapi nextjs no-code-scraping helm data-export automation data-engineering docker kubernetes web-server

2 sources

Member repositories

RepositoryRoleHealth v2
jaypyles/Scraperrmain10

For agents

markdown · JSON · MCP: product_card(name="jaypyles/Scraperr")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem