Skip to main content
在 Manus 中运行任何 Skill
一键导入

reviewing-agentic-safety

星标6
分支0
更新时间2026年7月22日 23:27

Reviews the action/tool surface of agent and tool-using systems — what the model is permitted to *do*, distinct from reviewing-llm-integration's review of the *model call*: tool least-privilege, approval gates and step/spend budgets on autonomous loops, tool metadata and MCP server descriptions as untrusted input (tool poisoning), confused-deputy and token-audience discipline (no token passthrough), inter-agent authentication, sandboxed code execution (no ambient credentials, egress allow-list), agent-memory hygiene, action audit trails, and the action-leg mitigations of the lethal trifecta. Grounded in OWASP's Top 10 for Agentic Applications (ASI01–ASI10) and the MCP security spec. Use when reviewing tool or function definitions exposed to a model, an MCP server or client, an autonomous or multi-agent loop, agent memory, or code that lets a model take actions. Skip when the change has no tools, agents, MCP, or autonomous loop — an ordinary model call with no action surface is reviewing-llm-integration's job.

安装

用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。

文件资源管理器
6 个文件
SKILL.md
readonly