Skip to main content
Manusで任意のスキルを実行
ワンクリックで

detect-tool-output-exfiltration-instructions

スター3
フォーク0
更新日2026年7月9日 16:16

Detect MCP tool-call responses that explicitly instruct an agent to exfiltrate conversation history, prompts, files, or secrets. Consumes native or OCSF Application Activity records from ingest-mcp-proxy-ocsf and emits OCSF 1.8 Detection Finding (class 2004) when a `tools/call` response contains high-confidence exfiltration language such as "send the conversation history", "upload local files", or "exfiltrate secrets". Use when the user mentions "tool-result exfiltration prompt injection", "response-layer data exfil instructions", or "MCP output says to upload secrets". Do NOT use for leaked credential material itself, tool descriptions, or semantic jailbreak claims.

インストール

Codex または Claude でインストール この Prompt をコピーして Codex、Claude、または他のアシスタントに貼り付けると、Skill ページを確認してインストールできます。

ファイルエクスプローラー
4 ファイル
SKILL.md
readonly