一键导入
judge
Use when uma tarefa de alto esforco ou dificuldade precisa de um juiz robusto e contraparte adversarial antes de concluir.
用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。
菜单
Use when uma tarefa de alto esforco ou dificuldade precisa de um juiz robusto e contraparte adversarial antes de concluir.
用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。
基于 SOC 职业分类
| name | judge |
| description | Use when uma tarefa de alto esforco ou dificuldade precisa de um juiz robusto e contraparte adversarial antes de concluir. |
Em tarefas de maior esforco/dificuldade (risco alto, decisao arquitetural, plano
complexo, mudanca sensivel), acione um juiz robusto para avaliar o artefato pelos
7 criterios da rubrica e produzir uma contraparte adversarial que melhore a
resposta. Nao e o fluxo comum — e uma escalada deliberada. Integra com a skill
adversarial-review.
scripts/judge.sh '<arquivo>' '<contexto/pedido>'.Use when delegating bounded subtasks to another agent or lighter model through Codex CLI, Hermes CLI, Antigravity CLI, or a local orchestration script to save tokens without losing control of scope and evidence.
Use when a task depends on library, framework, SDK, API, CLI, cloud service, version-specific behavior, setup instructions, migration guidance, or any external documentation that may be stale or mutable.
Use when debugging panes, errors, anomalies, regressions, flaky behavior, broken tests, incidents, unexpected UI behavior, or any failure where jumping straight to a fix could hide the real cause.
Use when a task changes UI, layout, frontend behavior, browser interaction, screenshots, visual regressions, E2E flows, accessibility-visible states, or anything that must be inspected in a browser.
Use when delivering a high-risk plan, architecture, audit, migration, security conclusion, refactor strategy, or any non-obvious technical recommendation that could be wrong in costly ways.
Use when choosing whether to answer briefly, write a technical plan, perform an audit, implement code, run incremental execution, or stop for confirmation based on ambiguity, risk, reversibility, and evidence needs.