Skip to main content
تشغيل أي مهارة في Manus
بنقرة واحدة

agent-audit-grade

النجوم٣
التفرعات١
آخر تحديث٢ يونيو ٢٠٢٦ في ٠٣:٠١

Second subagent in the agent-audit pipeline. Grades every assertion in evals-[n].json against actual_output using LLM judgment for semantic checks and tool calls for mechanical checks. Writes grading.json with per-assertion pass/fail and evidence, and updates evals-[n].json with a verdict per case. Use when agent-audit hands off "grade the evals".

التثبيت

التثبيت باستخدام Codex أو Claude انسخ هذا Prompt والصقه في Codex أو Claude أو مساعد آخر ليراجع صفحة Skill ويثبّتها لك.

SKILL.md
readonly