Skip to main content
在 Manus 中运行任何 Skill
一键导入

autorater-rubric

星标0
分支0
更新时间2026年5月23日 12:34

Use this skill when the user needs to build, calibrate, or critique an LLM-based autorater for UX, prose, agent output, or any subjective quality dimension. Covers rubric design, SxS vs SSE selection, anchor writing, the calibration loop, Cohen's kappa / Krippendorff's alpha agreement thresholds, and the anti-patterns that cause judge models to silently disagree with humans. Triggers on "build an autorater", "score this UX", "eval rubric", "LLM-as-judge", "calibrate a judge", "human-AI agreement", "judge model", "rubric design", "side-by-side eval", "SxS vs SSE", "inter-rater reliability", "kappa", "Krippendorff", "autorater calibration", "preference scoring", "scoring rubric".

安装

用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。

SKILL.md
readonly