Ross ROSS = Recommend OSS · open-source software intelligence for agents

meridianlabs-ai/inspect_petri

An alignment auditing agent capable of quickly exploring alignment hypothesis observed · 2026-08-28

github.com/meridianlabs-ai/inspect_petri · homepage · Python · MIT (permissive) observed · 2026-08-28

Health v2 · maintenance only

80/100

  • Activity 99
  • Release rhythm 85
  • Longevity 27
How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: 33
  • age_days: 379
  • days_rel: 21
  • days_push: 9
  • n_releases_24m: 4

Full methodology

Adoption not part of the score

1303 stars · 214 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

Inspect Petri is an alignment auditing agent that automatically probes language models for concerning behaviors like sycophancy, reward hacking, and eval awareness. It generates audit scenarios, orchestrates multi-turn conversations between an auditor and target model, and scores transcripts with a judge model.

Use cases

  • audit llm alignment automatically
  • test if my model is sycophantic
  • detect reward hacking in language models
  • red team an ai assistant
  • check if a model knows it is being evaluated
  • run multi-turn safety evaluations on llms
  • score llm transcripts for concerning behavior

When to choose

  • you need automated, hypothesis-driven alignment audits of language models
  • you want built-in conversation seeds and rubric-based judge scoring out of the box
  • you are doing AI safety red-teaming and want reproducible multi-turn audit pipelines

When to avoid

  • you need general-purpose LLM application evaluation rather than alignment-focused auditing
  • you want a simple benchmark suite without agent-driven scenario generation
  • you are not working with LLMs or AI safety evaluation at all

Facets

library · maturity active

agent-framework llm-inference testing benchmarking cli artificial-intelligence large-language-models machine-learning developer-tools python cli cross-platform alignment-auditing llm-red-teaming safety-evaluation reward-hacking-detection sycophancy-testing judge-model multi-turn-audit ai-safety ai-agents

3 sources

Member repositories

RepositoryRoleHealth v2
meridianlabs-ai/inspect_petrimain80

For agents

markdown · JSON · MCP: product_card(name="meridianlabs-ai/inspect_petri")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem