Skip to main content
在 Manus 中运行任何 Skill
一键导入

mimic-eval

星标4
分支0
更新时间2026年5月5日 22:01

Run Mimic-generated eval scenarios against a target agent skill (default `briefing-prep`) inside this Claude Code session. Reads `.mimic/exports/mimic-scenarios.json`, sub-agents one run per scenario via the Agent tool, scores responses with hybrid (strict substring + LLM-judge-on-miss + numeric-range) checks, and prints a scored table. Use when the user says any of "run the mimic eval", "score the briefing agent", "eval the brief against the facts", "/mimic-eval".

安装

用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。

SKILL.md
readonly