Ross ROSS = Recommend OSS · open-source software intelligence for agents

confident-ai/deepteam

DeepTeam is a framework to red team LLMs and AI agents. observed · 2026-08-28

github.com/confident-ai/deepteam · homepage · Python · Apache-2.0 (permissive) observed · 2026-08-28

Health v2 · maintenance only

67/100

  • Activity 98
  • Release rhythm 44
  • Longevity 39
How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.

  • gap_med: 86.0
  • age_days: 546
  • days_rel: 294
  • days_push: 12
  • n_releases_24m: 3

Full methodology

Adoption not part of the score

2623 stars · 426 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

DeepTeam is an open-source Python framework for red teaming LLM systems, AI agents, RAG pipelines, and chatbots. It simulates adversarial attacks such as jailbreaking and prompt injection to surface vulnerabilities, and maps risk assessments to frameworks like OWASP Top 10 for LLMs, NIST AI RMF, MITRE ATLAS, and the EU AI Act.

Use cases

  • red team my llm application for jailbreaks and prompt injection
  • test my ai agent for security vulnerabilities before production
  • run an OWASP Top 10 for LLMs risk assessment from python
  • detect PII leakage and bias in my chatbot
  • simulate multi-turn adversarial attacks against my RAG pipeline
  • check my llm app against NIST AI RMF or EU AI Act requirements
  • add guardrails to block unsafe prompts in production

When to choose

  • you need adversarial security testing for LLM apps, agents, or RAG systems
  • you want framework-aligned assessments (OWASP, NIST, MITRE ATLAS, EU AI Act) from a Python API or YAML CLI
  • you want a local, open-source alternative to commercial LLM pentesting tools
  • you already use DeepEval and want dedicated red teaming alongside evaluation

When to avoid

  • you need general LLM quality evaluation like correctness or faithfulness - use DeepEval instead
  • you need traditional network or web application penetration testing
  • you need a hosted, no-code-only security platform without local execution

Facets

framework · maturity active

security penetration-testing testing llm-inference agent-framework rag chatbot cli security large-language-models machine-learning developer-tools penetration-testing python cli cross-platform llm-red-teaming llm-safety llm-guardrails jailbreaking prompt-injection owasp-top-10 nist-ai-rmf mitre-atlas eu-ai-act vulnerability-scanning ai-security ai-agents

10 sources

Member repositories

RepositoryRoleHealth v2
confident-ai/deepteammain67

For agents

markdown · JSON · MCP: product_card(name="confident-ai/deepteam")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem