Open-Reasoner-Zero/Open-Reasoner-Zero
Official Repo for Open-Reasoner-Zero observed · 2026-08-28
Health v2 · maintenance only
31/100
- Activity 24
- Release rhythm 35
- Longevity 40
Flags: no_releases
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.
- gap_med: n/a
- age_days: 560
- days_rel: n/a
- days_push: 457
- n_releases_24m: 0
Adoption not part of the score
2099 stars · 120 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded
Open-Reasoner-Zero is an open-source implementation of large-scale reinforcement learning training for reasoning-oriented language models, released with code, hyperparameters, training data, and model weights. It reproduces and improves upon DeepSeek-R1-Zero-style RL training, achieving strong benchmark results with fewer training steps.
Use cases
- train a reasoning LLM with reinforcement learning from a base model
- reproduce DeepSeek-R1-Zero style RL training pipeline
- scale up RL training across model sizes from 0.5B to 32B
- improve math reasoning benchmarks like AIME2024 and MATH500
- research minimalist RL recipes for LLM reasoning
- download pretrained reasoning model weights for fine-tuning
When to choose
- you want an open, reproducible RL training pipeline for reasoning models
- you have GPU cluster resources to train or fine-tune large language models
- you need training data, hyperparameters, and weights released together for research
- you want to study scalability of RL on base models
When to avoid
- you only need inference of a reasoning model without training
- you lack multi-GPU infrastructure for large-scale RL training
- you need a general-purpose RL library unrelated to LLMs
- you want a plug-and-play chat model with no training involved
Facets
library · maturity active
reinforcement-learning llm-training machine-learning reinforcement-learning large-language-models deep-learning machine-learning python reasoning rl-training llm open-source-research deepseek-r1 model-training gpu linux
2 sources
- readme: https://github.com/Open-Reasoner-Zero/Open-Reasoner-Zero · fetched 2026-08-28 · a603172e5b6f
- homepage: https://yasminezhang.notion.site/Open-Reasoner-Zero-19e12cf72d418007b9cdebf44b0e7903 · fetched 2026-08-29 · 73a6ba54b760
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| Open-Reasoner-Zero/Open-Reasoner-Zero | main | 31 |
For agents
markdown · JSON · MCP: product_card(name="Open-Reasoner-Zero/Open-Reasoner-Zero")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem