openai/following-instructions-human-feedback resource
None observed · 2026-08-28
Health v2 · maintenance only
10/100
- Activity 0
- Release rhythm 35
- Longevity 100
Flags: no_releases archived no_license
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.
- gap_med: n/a
- age_days: 1682
- days_rel: n/a
- days_push: 1361
- n_releases_24m: 0
Adoption not part of the score
1259 stars · 149 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded
The official companion repository for OpenAI's InstructGPT paper on aligning language models with human intent via reinforcement learning from human feedback (RLHF). It contains the model card, evaluation samples, and labeling instructions rather than runnable training code.
Use cases
- understand how RLHF aligns language models with user intent
- read the InstructGPT model card and evaluation methodology
- study labeling instructions used for human feedback data collection
- compare GPT-3 and InstructGPT outputs on NLP benchmarks
- learn how supervised fine-tuning plus reward modeling reduces toxicity
- research alignment techniques for large language models
When to choose
- you want the authoritative artifacts and documentation behind the InstructGPT paper
- you are studying RLHF and model alignment from primary sources
- you need the official model card or labeling guidelines for citation or replication
When to avoid
- you need runnable RLHF training code - this repo ships no implementation
- you want the actual InstructGPT weights or datasets - they are not included
- you need a maintained library - the repo is a static paper companion
Facets
learning-resource · maturity maintenance
machine-learning llm-training reinforcement-learning nlp large-language-models machine-learning artificial-intelligence tutorials python rlhf instructgpt human-feedback model-alignment paper-companion model-card gpt-3 fine-tuning natural-language-processing gpu linux
1 source
- readme: https://github.com/openai/following-instructions-human-feedback · fetched 2026-08-28 · 8ba860f7ffca
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| openai/following-instructions-human-feedback | main | 10 |
For agents
markdown · JSON · MCP: product_card(name="openai/following-instructions-human-feedback")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem