Skip to main content
在 Manus 中运行任何 Skill
一键导入

eval-run

星标1
分支0
更新时间2026年4月30日 10:02

Use this skill whenever the user wants to execute a model evaluation run — testing a trained model on a dataset to measure metrics like mAP, accuracy, IoU, precision, recall, or per-class AP. Trigger for: launching eval runs (debug or production mode), running evaluation on a remote server, checking status of a running/crashed eval job, collecting results when eval finishes, forking a previous eval run with changed parameters (threshold, NMS, confidence, dataset split), and comparing metrics against a baseline. Also trigger for Chinese requests like "跑评估", "测一下", "跑一下eval", "对比baseline". This is the execution skill — not for initial config setup (use eval-init) or HTML report generation (use eval-report).

安装

用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。

SKILL.md
readonly