Skip to main content
在 Manus 中运行任何 Skill
一键导入

detect-tool-output-exfiltration-instructions

星标3
分支0
更新时间2026年7月9日 16:16

Detect MCP tool-call responses that explicitly instruct an agent to exfiltrate conversation history, prompts, files, or secrets. Consumes native or OCSF Application Activity records from ingest-mcp-proxy-ocsf and emits OCSF 1.8 Detection Finding (class 2004) when a `tools/call` response contains high-confidence exfiltration language such as "send the conversation history", "upload local files", or "exfiltrate secrets". Use when the user mentions "tool-result exfiltration prompt injection", "response-layer data exfil instructions", or "MCP output says to upload secrets". Do NOT use for leaked credential material itself, tool descriptions, or semantic jailbreak claims.

安装

用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。

文件资源管理器
4 个文件
SKILL.md
readonly