Ross ROSS = Recommend OSS · open-source software intelligence for agents

scrapfly/scrapfly-scrapers

Scalable Python web scraping scripts for +40 popular domains observed · 2026-09-03

github.com/scrapfly/scrapfly-scrapers · homepage · Python · NOASSERTION (other) observed · 2026-09-03

Health v2 · maintenance only

74/100

  • Activity 100
  • Release rhythm 35
  • Longevity 85

Flags: no_releases no_license

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 1199
  • days_rel: n/a
  • days_push: 0
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

1074 stars · 204 forks observed · 2026-09-03

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

A collection of educational Python web scraping scripts for over 40 popular domains such as Amazon, AliExpress, BestBuy, and Twitter, built on the Scrapfly web scraping API and its Python SDK. Each scraper comes with a companion tutorial, sample datasets, and uses an async stack of Scrapfly SDK, parsel, asyncio, JMESPath, and loguru.

Use cases

  • scrape product data and reviews from amazon
  • extract search results and product listings from aliexpress
  • scrape twitter posts without getting blocked
  • bypass anti-bot protection and captchas when crawling
  • learn how to build web scrapers for popular websites
  • collect e-commerce pricing data at scale
  • parse html and json responses from protected websites

When to choose

  • you need working, tested scraper examples for well-known sites like Amazon, AliExpress, or Twitter
  • you already use or plan to use the Scrapfly API and want reference implementations
  • you want to learn web scraping techniques through guided tutorials with sample datasets
  • you need anti-bot bypass, proxy rotation, and JS rendering handled by a managed service

When to avoid

  • you need a fully free, self-hosted scraping solution with no API dependency or usage costs
  • you want a general-purpose scraping framework like Scrapy rather than per-site scripts
  • your target website is not among the 40+ supported domains and you need custom scraping logic
  • you require a permissive open-source license for commercial redistribution, as the license is not a standard OSI-approved one

Facets

library · maturity active

web-scraping parser http-client workflow-automation crawlers web-development data-science e-commerce developer-tools python cli cross-platform scrapfly-api anti-bot-bypass captcha-bypass asyncio parsel jmespath loguru sample-datasets scrape-guides educational-scrapers proxy-rotation javascript-rendering e-commerce-scraping social-media-scraping twitter-scraper amazon-scraper aliexpress-scraper bestbuy-scraper spider-scripts data-extraction

10 sources

Member repositories

RepositoryRoleHealth v2
scrapfly/scrapfly-scrapersmain74

For agents

markdown · JSON · MCP: product_card(name="scrapfly/scrapfly-scrapers")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem