Skip to main content

coeval

Operate the CoEval medical LLM evaluation repository: set up the toolchain, compose Hydra configs, run or monitor evaluations, configure OpenAI-compatible candidate and judge models, and audit or compare evaluation artifacts. Use when working in this repository or when asked about CoEval runs, HealthBench, passthrough clients, retries, Hydra overrides, or evaluation_outputs. Do not use for generic evaluation work outside CoEval.

Jump to install

Source facts

Repository
lunit-io/CoEval
Last source activity
August 18, 2026 at 07:46
Detected SKILL.md language
English
Stars
5
Forks
2

Install options

The review-first prompt is selected by default. You can switch to a direct command or download a local copy.

Review the source files

Read SKILL.md and any companion files shown by SkillsMP before deciding whether to install.