WeiboAI/VibeThinker resource
Tiny Model, Big Logic: Diversity-Driven Optimization Elicits Large-Model Reasoning Ability in VibeThinker-1.5B observed · 2026-08-28
Health v2 · maintenance only
60/100
- Activity 97
- Release rhythm 35
- Longevity 21
Flags: no_releases
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.
- gap_med: n/a
- age_days: 302
- days_rel: n/a
- days_push: 19
- n_releases_24m: 0
Adoption not part of the score
1561 stars · 116 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded
VibeThinker is a family of small dense reasoning language models (1.5B and 3B parameters) from WeiboAI, trained with a Spectrum-to-Signal post-training pipeline to achieve frontier-level math and coding reasoning. The repository hosts documentation, papers, and links to model weights on Hugging Face and ModelScope.
Use cases
- run a small reasoning model for math olympiad problems
- download a 1.5B LLM that rivals much larger reasoning models
- study how diversity-driven post-training elicits reasoning in tiny models
- benchmark small models on AIME and LiveCodeBench
- deploy a low-cost reasoning LLM locally
- research small-model reinforcement learning pipelines
When to choose
- you need strong math/competitive-programming reasoning from a tiny model that fits on modest hardware
- you want to study or reproduce the Spectrum-to-Signal post-training approach
- you need open MIT-licensed reasoning model weights
When to avoid
- you need general-purpose chat, long-context, or multimodal capabilities
- you require frontier-scale knowledge breadth rather than verifiable reasoning tasks
- you want a ready-made inference server or application rather than model weights
Facets
learning-resource · maturity active
machine-learning llm-training llm-inference large-language-models machine-learning artificial-intelligence python reasoning-models small-language-models post-training reinforcement-learning math-reasoning model-weights huggingface gpu
1 source
- readme: https://github.com/WeiboAI/VibeThinker · fetched 2026-08-28 · 56c96955c3e7
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| WeiboAI/VibeThinker | main | 60 |
For agents
markdown · JSON · MCP: product_card(name="WeiboAI/VibeThinker")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem