Skip to main content

output-eval-validate-judge

Validate LLM judges against human labels using TPR/TNR metrics and train/dev/test splits. Use after writing a judge prompt to verify it agrees with human judgment.

Jump to install

Source facts

Repository
growthxai/output
Last source activity
June 23, 2026 at 19:41
Detected SKILL.md language
English
Stars
434
Forks
12

Install options

The review-first prompt is selected by default. You can switch to a direct command or download a local copy.

Review the source files

Read SKILL.md and any companion files shown by SkillsMP before deciding whether to install.