dataabc/weibo-crawler
新浪微博爬虫,用python爬取新浪微博数据,并下载微博图片和微博视频 observed · 2026-08-28
Health v2 · maintenance only
74/100
- Activity 93
- Release rhythm 35
- Longevity 100
Flags: no_releases no_license
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.
- gap_med: n/a
- age_days: 2626
- days_rel: n/a
- days_push: 42
- n_releases_24m: 0
Adoption not part of the score
4625 stars · 912 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded
A Python crawler for Sina Weibo that scrapes user profiles and posts, exporting data to CSV, JSON, MySQL, MongoDB, or SQLite, and optionally downloading images, videos, comments, and reposts. It supports incremental crawling on a schedule and can run via Docker or as an API service.
Use cases
- scrape a Weibo user's posts and profile data
- download images and videos from Weibo accounts
- export Weibo posts to CSV or JSON
- store Weibo data in MySQL or MongoDB
- incrementally crawl new Weibo posts on a schedule
- collect Weibo comments and reposts for analysis
When to choose
- you need structured Weibo data for research or analysis
- you want media (images/videos) downloaded alongside post metadata
- you need flexible output formats including databases
- you want scheduled incremental crawling of Weibo accounts
When to avoid
- you need to crawl platforms other than Sina Weibo
- you need a GUI rather than config-file-driven CLI usage
- your use case requires an officially supported API with guaranteed compliance
Facets
cli-tool · maturity active
web-scraping data-generation etl crawlers social-media python cli cross-platform weibo sina-weibo scraper media-download csv mysql mongodb sqlite data-engineering docker
1 source
- readme: https://github.com/dataabc/weibo-crawler · fetched 2026-08-28 · f79ad3d47713
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| dataabc/weibo-crawler | main | 74 |
For agents
markdown · JSON · MCP: product_card(name="dataabc/weibo-crawler")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem