Run local adversarial attack passes against agents, RAG pipelines, and chatbots to surface concrete failure classes before production rollout.
Scanned 6/2/2026
Install to Claude Code
npx -y skills add agentskillexchange/skills --skill red-team-agent-workflows-for-jailbreaks-prompt-injection-and-policy-failures-with-deepteam --agent claude-codeInstalls into .claude/skills of the current project.
Are you the author of Red Team Agent Workflows For Jailbreaks Prompt Injection And Policy Failures With Deepteam?
Add the live security badge to your README — it updates automatically with every re-scan.
[](https://www.skillsdirectory.com/skills/agentskillexchange-red-team-agent-workflows-for-jailbreaks-prompt-inj)More formats (shields.io, HTML) on the badges page.
---
name: "Red-team agent workflows for jailbreaks, prompt injection, and policy failures with DeepTeam"
slug: "red-team-agent-workflows-for-jailbreaks-prompt-injection-and-policy-failures-with-deepteam"
description: "Run local adversarial attack passes against agents, RAG pipelines, and chatbots to surface concrete failure classes before production rollout."
github_stars: 1566
verification: "security_reviewed"
source: "https://github.com/confident-ai/deepteam"
author: "Confident AI"
publisher_type: "organization"
category: "Security & Verification"
framework: "Multi-Framework"
tool_ecosystem:
github_repo: "confident-ai/deepteam"
github_stars: 1566
---
# Red-team agent workflows for jailbreaks, prompt injection, and policy failures with DeepTeam
Run local adversarial attack passes against agents, RAG pipelines, and chatbots to surface concrete failure classes before production rollout.
## Prerequisites
Python environment, local or configured LLM access for chosen attacks
## Installation
Use the upstream install or setup path that matches your environment:
- pip install -U deepteam
Requirements and caveats from upstream:
- 🔗 Run red teaming from the **CLI** with YAML configs, or programmatically in Python.
- DeepTeam does not require you to define what LLM system you are red teaming — because neither will malicious users. All you need to do is install deepteam, define a model_callback, and you're good to go.
- python
Basic usage or getting-started notes:
- <a href="#-quickstart">Getting Started</a> |
- 📐 50+ ready-to-use [vulnerabilities](https://www.trydeepteam.com/docs/red-teaming-vulnerabilities) (all with explanations) powered by **ANY** LLM of your choice. Each vulnerability uses LLM-as-a-Judge metrics that run...
- ## Red Team Your First LLM
- Source: https://github.com/confident-ai/deepteam
- Extracted from upstream docs: https://raw.githubusercontent.com/confident-ai/deepteam/HEAD/README.md
## Documentation
- https://github.com/confident-ai/deepteam
## Source
- [Agent Skill Exchange](https://agentskillexchange.com/skills/red-team-agent-workflows-for-jailbreaks-prompt-injection-and-policy-failures-with-deepteam/)
Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.
No comments yet. Be the first to comment!