Skip to main content
Run any Skill in Manus
with one click

autorater-rubric

Stars0
Forks0
UpdatedMay 23, 2026 at 12:34

Use this skill when the user needs to build, calibrate, or critique an LLM-based autorater for UX, prose, agent output, or any subjective quality dimension. Covers rubric design, SxS vs SSE selection, anchor writing, the calibration loop, Cohen's kappa / Krippendorff's alpha agreement thresholds, and the anti-patterns that cause judge models to silently disagree with humans. Triggers on "build an autorater", "score this UX", "eval rubric", "LLM-as-judge", "calibrate a judge", "human-AI agreement", "judge model", "rubric design", "side-by-side eval", "SxS vs SSE", "inter-rater reliability", "kappa", "Krippendorff", "autorater calibration", "preference scoring", "scoring rubric".

Installation

Install with Codex or Claude Copy this prompt, paste it into Codex, Claude, or another assistant, and let it review the skill page and install it for you.

SKILL.md
readonly