بنقرة واحدة
evaluation-judge
Scores competing AI, colens, or prompt outputs against a declared rubric with evidence.
التثبيت باستخدام Codex أو Claude انسخ هذا Prompt والصقه في Codex أو Claude أو مساعد آخر ليراجع صفحة Skill ويثبّتها لك.
القائمة
Scores competing AI, colens, or prompt outputs against a declared rubric with evidence.
التثبيت باستخدام Codex أو Claude انسخ هذا Prompt والصقه في Codex أو Claude أو مساعد آخر ليراجع صفحة Skill ويثبّتها لك.
استنادا إلى تصنيف SOC المهني
Creator-side battle comparing the blog outline and YouTube script for the same topic.
Battle template comparing Claude and OpenAI on the same code review task using a four-axis rubric.
Compares single-lens and workflow-based finance report explanations. This is not financial advice.
Compares two models on creator planning using the YouTube content workflow.
Battle template comparing AI legal-review outputs on identical contract text. Analysis only — NOT legal advice.
Compares legal-adjacent document review outputs. This is not legal advice and must be reviewed by a qualified lawyer.
| name | evaluation-judge |
| description | Scores competing AI, colens, or prompt outputs against a declared rubric with evidence. |
Use this lenser when a battle or comparison needs a consistent rubric-based judgment.
The lenser may judge generated outputs. It must not hide uncertainty or choose a winner without evidence.
Return a score table, short rationale per criterion, winner, and residual uncertainty.