Skip to main content

maragudk/evals-skills

SkillsMP는 maragudk/evals-skills에서 4개의 skill을 수집했습니다. skill을 열어 소스와 세부 정보를 확인하세요.

최근 기록된 소스 활동
SkillsMP 카탈로그 업데이트
수집된 skills
4
GitHub 스타
12
GitHub 포크
0

이 저장소의 skills

수집된 skill 4개 중 4개를 표시합니다.

직업 분류
소프트웨어 개발자
설명

Generate a custom trace annotation web app for open coding during LLM error analysis. Use when the user wants to review LLM traces, annotate failures with freeform comments, and do first-pass qualitative labeling (open coding). Also use when the user mentions…

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

Build a structured taxonomy of failure modes from open-coded trace annotations. Use this skill whenever the user has freeform annotations from reviewing LLM traces and wants to cluster them into a coherent, non-overlapping set of binary failure categories…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Use this skill when crafting, reviewing, or improving prompts for LLM pipelines — including task prompts, system prompts, and LLM-as-Judge prompts. Triggers include: requests to write or refine a prompt, diagnose why an LLM produces inconsistent or incorrect…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 품질 보증 분석가·테스터
설명

Build, validate, and deploy LLM-as-Judge evaluators for automated quality assessment of LLM pipeline outputs. Use this skill whenever the user wants to: create an automated evaluator for subjective or nuanced failure modes, write a judge prompt for Pass/Fail…

원문 언어: 영어

업데이트
수집된 skill 4개 중 4개를 표시합니다.