Ross ROSS = Recommend OSS · open-source software intelligence for agents

openai/following-instructions-human-feedback resource

None observed · 2026-08-28

github.com/openai/following-instructions-human-feedback · archived observed · 2026-08-28

Health v2 · maintenance only

10/100

  • Activity 0
  • Release rhythm 35
  • Longevity 100

Flags: no_releases archived no_license

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 1682
  • days_rel: n/a
  • days_push: 1361
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

1259 stars · 149 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

The official companion repository for OpenAI's InstructGPT paper on aligning language models with human intent via reinforcement learning from human feedback (RLHF). It contains the model card, evaluation samples, and labeling instructions rather than runnable training code.

Use cases

  • understand how RLHF aligns language models with user intent
  • read the InstructGPT model card and evaluation methodology
  • study labeling instructions used for human feedback data collection
  • compare GPT-3 and InstructGPT outputs on NLP benchmarks
  • learn how supervised fine-tuning plus reward modeling reduces toxicity
  • research alignment techniques for large language models

When to choose

  • you want the authoritative artifacts and documentation behind the InstructGPT paper
  • you are studying RLHF and model alignment from primary sources
  • you need the official model card or labeling guidelines for citation or replication

When to avoid

  • you need runnable RLHF training code - this repo ships no implementation
  • you want the actual InstructGPT weights or datasets - they are not included
  • you need a maintained library - the repo is a static paper companion

Facets

learning-resource · maturity maintenance

machine-learning llm-training reinforcement-learning nlp large-language-models machine-learning artificial-intelligence tutorials python rlhf instructgpt human-feedback model-alignment paper-companion model-card gpt-3 fine-tuning natural-language-processing gpu linux

1 source

Member repositories

RepositoryRoleHealth v2
openai/following-instructions-human-feedbackmain10

For agents

markdown · JSON · MCP: product_card(name="openai/following-instructions-human-feedback")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem