Ross ROSS = Recommend OSS · open-source software intelligence for agents

yhangf/PythonCrawler resource

:heartpulse:用python编写的爬虫项目集合 observed · 2026-08-28

github.com/yhangf/PythonCrawler · Python · MIT (permissive) observed · 2026-08-28

Health v2 · maintenance only

69/100

  • Activity 82
  • Release rhythm 35
  • Longevity 100

Flags: no_releases

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 3703
  • days_rel: n/a
  • days_push: 113
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

1820 stars · 493 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

A collection of Python web crawler scripts covering tasks like scraping images from Baidu, job postings, JD product data, and GitHub trending projects. It is intended as a learning resource for web scraping techniques.

Use cases

  • learn web scraping with python
  • scrape images from a website
  • crawl job postings by keyword
  • scrape jd product data
  • download all images from a webpage
  • example python spider scripts

When to choose

  • you want ready-made example scripts to learn scraping patterns
  • you need a quick starting point for a simple crawler task

When to avoid

  • you need a production-grade, maintained scraping framework
  • you need scalable distributed crawling with scheduling

Facets

learning-resource · maturity active

web-scraping developer-tools crawlers developer-tools tutorials python cli spiders scraping-examples python3 educational

1 source

Member repositories

RepositoryRoleHealth v2
yhangf/PythonCrawlermain69

For agents

markdown · JSON · MCP: product_card(name="yhangf/PythonCrawler")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem