Skip to main content
تشغيل أي مهارة في Manus
بنقرة واحدة

agent-audit-benchmark

النجوم٣
التفرعات١
آخر تحديث٢ يونيو ٢٠٢٦ في ٠٣:٠١

Fifth subagent in the agent-audit pipeline. Reads evals-[n].json and grading.json from the run dir, aggregates timing across all evals, computes pass rate and token stats, and writes timing.json and benchmark.json. Handles missing or partial data gracefully — writes what it can and flags gaps. Use when agent-audit hands off "compute the benchmark".

التثبيت

التثبيت باستخدام Codex أو Claude انسخ هذا Prompt والصقه في Codex أو Claude أو مساعد آخر ليراجع صفحة Skill ويثبّتها لك.

SKILL.md
readonly