Skip to content
appsgit

Redteam

Redteam is an agent skill (a SKILL.md file) from agentscope-ai/OpenJudge. Agents load it when the user wants to test their LLM/agent application for safety and security vulnerabilities — jailbreaks, prompt injection, PII extraction, harmful content generation, or evaluator gaming. It works with Claude Code and has 867 GitHub stars.

github.com/agentscope-ai/OpenJudge/skills/eval_pipeline/07-redteam (opens in a new tab)

Add this skill

Claude

Claude Code loads skills from ~/.claude/skills/ (all projects) or .claude/skills/ (one project):

git clone --depth 1 https://github.com/agentscope-ai/OpenJudge.git
cp -r OpenJudge/skills/eval_pipeline/07-redteam ~/.claude/skills/redteam   # personal, or .claude/skills in a project

In the Claude apps, zip the redteam folder and upload it under Customize > Skills > + > Upload a skill (code execution must be on).

ChatGPT / Codex

Codex reads skills from .agents/skills/ in a repo or ~/.agents/skills/ for every project:

git clone --depth 1 https://github.com/agentscope-ai/OpenJudge.git
cp -r OpenJudge/skills/eval_pipeline/07-redteam .agents/skills/redteam   # repo; ~/.agents/skills for all projects

Standalone skills also load in the ChatGPT desktop app.

Cursor

Cursor loads skills from .cursor/skills/ (or ~/.cursor/skills/) and also reads .claude/skills/:

git clone --depth 1 https://github.com/agentscope-ai/OpenJudge.git
cp -r OpenJudge/skills/eval_pipeline/07-redteam .cursor/skills/redteam   # project; ~/.cursor/skills for all projects

Source (checked Oct 7, 2026): code.claude.com/docs/en/skills (opens in a new tab), support.claude.com/en/articles/12512180-using-skills-in-claude (opens in a new tab), learn.chatgpt.com/docs/build-skills (opens in a new tab), cursor.com/docs/context/skills (opens in a new tab)

What this skill does

Use when the user wants to test their LLM/agent application for safety and security vulnerabilities — jailbreaks, prompt injection, PII extraction, harmful content generation, or evaluator gaming. Also use when the user mentions security testing, adversarial testing, red teaming, safety evaluation, ASR (Attack Success Rate), or "is my app safe to deploy." Outputs ASR paired with over-refusal rate and an audit document. Test your application's safety boundaries systematically.

When it triggers

  • Use when the user wants to test their LLM/agent application for safety and security vulnerabilities — jailbreaks, prompt injection, PII extraction, harmful content generation, or evaluator gaming.
  • use when the user mentions security testing, adversarial testing, red teaming, safety evaluation, ASR (Attack Success Rate), or "is my app safe to deploy.
  • "is my app safe to deploy."

FAQ

Redteam FAQ

Still curious? Email info@appsgit.com.

What is the Redteam skill?

Redteam is an agent skill (a SKILL.md file) from agentscope-ai/OpenJudge. Agents load it when the user wants to test their LLM/agent application for safety and security vulnerabilities — jailbreaks, prompt injection, PII extraction, harmful content generation, or evaluator gaming. It works with Claude Code and has 867 GitHub stars. Its SKILL.md lives at github.com/agentscope-ai/OpenJudge/skills/eval_pipeline/07-redteam.

How do I install the Redteam skill?

Copy the redteam folder (the one containing SKILL.md) into ~/.claude/skills/ for Claude Code, .agents/skills/ for Codex or .cursor/skills/ for Cursor. The agent picks it up automatically when a task matches its description.

Is the Redteam skill free?

Yes. The repository is open source under the Apache-2.0 license.

Is Redteam maintained?

The repository's most recent commit was on Sep 11, 2026. Its latest release is v0.2.2. appsgit only lists skills from repositories with a commit in the last six months.