# scrapfly/scrapfly-scrapers

Scalable Python web scraping scripts for +40 popular domains

Repository: https://github.com/scrapfly/scrapfly-scrapers
Canonical: https://ross.abutalabs.com/products/scrapfly-scrapers
Homepage: https://scrapfly.io
Language: Python
License: NOASSERTION
License Family: other
Topics: crawling, python, crawler, scraping, web-scraping, web-scraping-python, antibot, automation, captcha-bypass, crawling-python, datascraping, proxies, python-scraper, scraper, scraping-python, spider, twitter-scraper, web-crawler, webscraper, webscraping
Last push: 2026-09-02T21:25:46+00:00

## Health v2 (maintenance only)
Score: 74/100 (v2, computed 2026-09-03T02:20:16.233290+00:00)
- activity 100, release rhythm 35, longevity 85
- inputs: {"age_days": 1199, "days_push": 0, "days_rel": null, "gap_med": null, "n_releases_24m": 0}
- flags: no_releases, no_license
- formula: round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10)

## Adoption (not part of the score)
Stars 1074, forks 204 (observed 2026-09-03T02:15:17.719871+00:00)

## What it is
A collection of educational Python web scraping scripts for over 40 popular domains such as Amazon, AliExpress, BestBuy, and Twitter, built on the Scrapfly web scraping API and its Python SDK. Each scraper comes with a companion tutorial, sample datasets, and uses an async stack of Scrapfly SDK, parsel, asyncio, JMESPath, and loguru.

## Use cases
- scrape product data and reviews from amazon
- extract search results and product listings from aliexpress
- scrape twitter posts without getting blocked
- bypass anti-bot protection and captchas when crawling
- learn how to build web scrapers for popular websites
- collect e-commerce pricing data at scale
- parse html and json responses from protected websites

## When to choose
- you need working, tested scraper examples for well-known sites like Amazon, AliExpress, or Twitter
- you already use or plan to use the Scrapfly API and want reference implementations
- you want to learn web scraping techniques through guided tutorials with sample datasets
- you need anti-bot bypass, proxy rotation, and JS rendering handled by a managed service

## When to avoid
- you need a fully free, self-hosted scraping solution with no API dependency or usage costs
- you want a general-purpose scraping framework like Scrapy rather than per-site scripts
- your target website is not among the 40+ supported domains and you need custom scraping logic
- you require a permissive open-source license for commercial redistribution, as the license is not a standard OSI-approved one

## Facets
- artifact type: library
- maturity: active
- function: web-scraping, parser, http-client, workflow-automation
- domain: crawlers, web-development, data-science, e-commerce, developer-tools
- platform: python, cli, cross-platform
- tags: scrapfly-api, anti-bot-bypass, captcha-bypass, asyncio, parsel, jmespath, loguru, sample-datasets, scrape-guides, educational-scrapers, proxy-rotation, javascript-rendering, e-commerce-scraping, social-media-scraping, twitter-scraper, amazon-scraper, aliexpress-scraper, bestbuy-scraper, spider-scripts, data-extraction

## Member repositories
- scrapfly/scrapfly-scrapers (main) score 74

## Provenance
- Observed fields: from GitHub, fetched 2026-09-03T02:15:17.719871+00:00.
- Health v2: computed from the inputs above; adoption is never an input.
- Inferred fields (summary, facets, guidance): AI-extracted, prompt v1, taxonomy v1, on 2026-08-30T06:55:02.572826+00:00, confidence not recorded.
  - readme: https://github.com/scrapfly/scrapfly-scrapers (fetched 2026-09-03T02:15:17.719871+00:00, sha 80ff50a5c2b6)
  - homepage: https://scrapfly.io (fetched 2026-08-29T12:57:26.744112+00:00, sha c8172d9ff46c)
  - site_page: https://scrapfly.io/docs (fetched 2026-08-29T12:57:26.749042+00:00, sha 84b755dc975e)
  - site_page: https://scrapfly.io/docs/sdk/python (fetched 2026-08-29T12:57:26.750671+00:00, sha bec935df129a)
  - site_page: https://scrapfly.io/docs/sdk/typescript (fetched 2026-08-29T12:57:26.752851+00:00, sha 6cccbd692b94)
  - site_page: https://scrapfly.io/docs/sdk/golang (fetched 2026-08-29T12:57:26.754870+00:00, sha fa16a708674e)
  - site_page: https://scrapfly.io/docs/sdk/rust (fetched 2026-08-29T12:57:26.756877+00:00, sha a8b4856f656e)
  - site_page: https://scrapfly.io/docs/sdk/scrapy (fetched 2026-08-29T12:57:26.759087+00:00, sha 73c515dfa723)
  - site_page: https://scrapfly.io/docs/cli (fetched 2026-08-29T12:57:26.760725+00:00, sha 6f8e539e868a)
  - site_page: https://scrapfly.io/pricing (fetched 2026-08-29T12:57:26.747048+00:00, sha d281e3e44f92)
- Data as of 2026-08-30T08:39:29.467469+00:00.
