Ross ROSS = Recommend OSS · open-source software intelligence for agents

JailbrokenAI/wallbreaker

None observed · 2026-08-28

github.com/JailbrokenAI/wallbreaker · Python · AGPL-3.0 (copyleft) observed · 2026-08-28

Health v2 · maintenance only

57/100

  • Activity 97
  • Release rhythm 35
  • Longevity 4

Flags: no_releases young

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 68
  • days_rel: n/a
  • days_push: 22
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

1310 stars · 219 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

Wallbreaker is a Claude-Code-style terminal harness for red-teaming LLMs, driving an autonomous agent loop that runs jailbreak attacks (PAIR/TAP, Crescendo, best-of-N) against configurable OpenAI-/Anthropic-compatible backends. It bundles transform engines (Parseltongue), a jailbreak library, HarmBench benchmarks, an LLM judge, and reliability validation for measuring real bypass success rates.

Use cases

  • red-team an LLM for jailbreak vulnerabilities
  • run automated PAIR or TAP attack loops against a model
  • measure real jailbreak success rate with repeated validation
  • test system prompt robustness with HarmBench behaviors
  • generate adversarial prompt transforms like encodings and homoglyphs
  • author persona-based system prompt jailbreaks
  • run a universal system prompt extraction sweep

When to choose

  • you are doing authorized LLM security testing and want an agentic attack harness
  • you need standardized benchmarks (HarmBench) instead of hand-picked prompts
  • you want a configurable backend across OpenRouter, Anthropic, OpenAI-compatible, or local APIs
  • you need transform-based evasion like Parseltongue encodings and steganography

When to avoid

  • you want to jailbreak models without authorization - this is for authorized testing only
  • you need a general-purpose coding assistant rather than a red-team tool
  • you need a GUI or web interface - it is terminal-only
  • your use case is defensive-only filtering with no adversarial evaluation

Facets

cli-tool · maturity active

security llm-inference agent-framework mcp cli testing prompt-engineering security artificial-intelligence large-language-models penetration-testing developer-tools cli python cross-platform red-teaming jailbreak llm-security adversarial-attacks harmbench pair-tap crescendo openrouter anthropic-api openai-api authorized-testing command-line

1 source

Member repositories

RepositoryRoleHealth v2
JailbrokenAI/wallbreakermain57

For agents

markdown · JSON · MCP: product_card(name="JailbrokenAI/wallbreaker")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem