Skip to main content

coeval

Operate the CoEval medical LLM evaluation repository: set up the toolchain, compose Hydra configs, run or monitor evaluations, configure OpenAI-compatible candidate and judge models, and audit or compare evaluation artifacts. Use when working in this repository or when asked about CoEval runs, HealthBench, passthrough clients, retries, Hydra overrides, or evaluation_outputs. Do not use for generic evaluation work outside CoEval.

跳到安装

来源信息

仓库
lunit-io/CoEval
最近来源活动
2026年8月18日 07:46
检测到的 SKILL.md 语言
英语
星标
5
分支
2

安装方式

默认使用会先检查来源的 Prompt;你也可以切换为直接命令,或下载本地副本。

检查来源文件

决定是否安装前,请先阅读 SKILL.md,以及 SkillsMP 当前展示的配套文件。