Skip to main content

ai-evals-course/evals-skills

SkillsMP는 ai-evals-course/evals-skills에서 4개의 skill을 수집했습니다. skill을 열어 소스와 세부 정보를 확인하세요.

최근 기록된 소스 활동
SkillsMP 카탈로그 업데이트
수집된 skills
4
GitHub 스타
451
GitHub 포크
36

이 저장소의 skills

수집된 skill 4개 중 4개를 표시합니다.

직업 분류
소프트웨어 품질 보증 분석가·테스터
설명

Entry point for evals. Use when the user asks for help with evals, does not know where to begin, or asks for something no other skill in this plugin matches. Do NOT use when a more specific skill in this plugin already matches; load that skill directly.

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

Run error analysis on a dataset. Build a review UI, select diverse samples, monitor annotations, and organize failure modes.

원문 언어: 영어

업데이트
직업 분류
웹 개발자
설명

Build a custom browser-based annotation interface tailored to your data for reviewing LLM traces and collecting structured feedback. Use when you need to build an annotation tool, review traces, or collect human labels.

원문 언어: 영어

업데이트
직업 분류
소프트웨어 품질 보증 분석가·테스터
설명

Audit an LLM eval pipeline and surface problems: missing error analysis, unvalidated judges, vanity metrics, etc. Use when inheriting an eval system, when unsure whether evals are trustworthy, or as a starting point when no eval infrastructure exists. Do NOT…

원문 언어: 영어

업데이트
수집된 skill 4개 중 4개를 표시합니다.