Skip to main content

braintrust-elicit-eval-criteria

Extract evaluation criteria out of domain experts and real user desires, and capture them as reusable evaluation assets before any labeling or scoring begins — construct facets, anchored exemplars, adversarial traps, scoring guidance, audit rules, and the signals that reveal what users actually want. Use when nobody can say what "good" means, when a rubric does not exist yet, when expert knowledge lives only in reviewers' heads, or when validating that an eval reflects user desires rather than team assumptions. Do not use to run the labeling workflow or to compare a scorer against finished labels.

跳到安装

来源信息

仓库
braintrustdata/eval-library
最近来源活动
2026年8月17日 20:49
检测到的 SKILL.md 语言
英语
星标
12
分支
2

安装方式

默认使用会先检查来源的 Prompt;你也可以切换为直接命令,或下载本地副本。

检查来源文件

决定是否安装前,请先阅读 SKILL.md,以及 SkillsMP 当前展示的配套文件。