Skip to main content
在 Manus 中运行任何 Skill
一键导入

llm-redteam-hunter

星标15
分支7
更新时间2026年6月28日 16:46

Automated red-teaming of the organization's OWN LLM / inference endpoints using garak and PyRIT. Runs probe batteries for prompt injection, jailbreak / guardrail bypass, harmful-content generation, training-data / system-prompt leakage, and insecure output handling, then normalizes the scanner output into the canonical finding schema mapped to the OWASP LLM Top 10 (LLM01 Prompt Injection, LLM02 Insecure Output, LLM06 Sensitive Info Disclosure, LLM07 System-Prompt Leakage). Complements (does not overlap) the manual /redteam-ai command and security-analyst OBLITERATUS work by adding automated, repeatable scanning. Use when an in-scope endpoint embeds or exposes an LLM. Requires .claude/security-scope.yaml with the target listed under llm_endpoints and llm_redteam: approved. Defensive testing of your own endpoints only.

安装

用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。

SKILL.md
readonly