Skip to main content
Manusで任意のスキルを実行
ワンクリックで

autorater-rubric

スター0
フォーク0
更新日2026年5月23日 12:34

Use this skill when the user needs to build, calibrate, or critique an LLM-based autorater for UX, prose, agent output, or any subjective quality dimension. Covers rubric design, SxS vs SSE selection, anchor writing, the calibration loop, Cohen's kappa / Krippendorff's alpha agreement thresholds, and the anti-patterns that cause judge models to silently disagree with humans. Triggers on "build an autorater", "score this UX", "eval rubric", "LLM-as-judge", "calibrate a judge", "human-AI agreement", "judge model", "rubric design", "side-by-side eval", "SxS vs SSE", "inter-rater reliability", "kappa", "Krippendorff", "autorater calibration", "preference scoring", "scoring rubric".

インストール

Codex または Claude でインストール この Prompt をコピーして Codex、Claude、または他のアシスタントに貼り付けると、Skill ページを確認してインストールできます。

SKILL.md
readonly