Skip to main content

noisy-but-valid-robust

Statistically certify LLM safety/quality using imperfect LLM judges with guaranteed Type-I error control. Implements the "Noisy but Valid" hypothesis testing framework: calibrate a judge's TPR/FPR on a small human-labeled set, then run a variance-corrected test on a large judge-labeled dataset. Use when: "certify my model's failure rate", "validate LLM safety with an LLM judge", "statistical test with noisy labels", "is my model below the safety threshold", "evaluate LLM with imperfect judge", "calibrate judge accuracy and run hypothesis test".

Jump to install

Source facts

Repository
ndpvt-web/arxiv-claude-skills
Last source activity
February 13, 2026 at 08:37
Detected SKILL.md language
English
Stars
14
Forks
3

Install options

The review-first prompt is selected by default. You can switch to a direct command or download a local copy.

Review the source files

Read SKILL.md and any companion files shown by SkillsMP before deciding whether to install.