nghuyong/WeiboSpider
持续维护的新浪微博采集工具🚀🚀🚀 observed · 2026-08-28
Health v2 · maintenance only
73/100
- Activity 90
- Release rhythm 35
- Longevity 100
Flags: no_releases
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.
- gap_med: n/a
- age_days: 3230
- days_rel: n/a
- days_push: 64
- n_releases_24m: 0
Adoption not part of the score
4109 stars · 836 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded
A continuously maintained Python web scraping tool for Sina Weibo built on Scrapy and the new weibo.com API. It collects user profiles, posts, followers, follows, reposts, comments, and keyword search results into JSONL files.
Use cases
- scrape weibo user profiles and posts
- collect weibo comments and reposts for research
- search weibo posts by keyword and save results
- gather weibo follower and following lists
- build a Chinese social media dataset for NLP
When to choose
- you need structured Weibo data with rich fields from the current weibo.com API
- you want a small, readable Scrapy codebase you can customize quickly
- you need multiple collection modes (users, posts, comments, search) in one tool
When to avoid
- you need to scrape platforms other than Weibo
- you cannot provide a valid logged-in weibo.com cookie
- you need a no-code or GUI scraping solution
Facets
cli-tool · maturity active
web-scraping data-generation etl crawlers social-media python cli cross-platform weibo scrapy sina-weibo social-media-scraping l-output data-engineering natural-language-processing
1 source
- readme: https://github.com/nghuyong/WeiboSpider · fetched 2026-08-28 · 96d21053bd0d
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| nghuyong/WeiboSpider | main | 73 |
For agents
markdown · JSON · MCP: product_card(name="nghuyong/WeiboSpider")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem