Skip to main content

rhoai-model-evaluation

Guide RHOAI evaluation workflows once active evaluation content exists; during the reimplementation, use this skill to rebuild EvalHub, RAG evaluation, KFP/MLflow evidence, RAGAS patterns, and standard model benchmarking workflows from legacy references. Use when the user asks to run evaluation, evaluate a model, benchmark model performance, check RAG quality, compare pre-RAG vs post-RAG answers, run LM-Eval, create an LMEvalJob, interpret eval results, add new test questions, or modify the judge prompt. Also use when eval pipelines fail, LMEvalJob pods are stuck, or evaluation reports show unexpected scores. Do NOT use for official product EvalHub, LM-Eval, LMEvalJob, or automated risk assessment workflows from the Red Hat evaluation guide (use rhoai-evaluation), product AutoRAG dashboard optimization runs, leaderboard review, or generated notebooks (use rhoai-autorag), MLflow platform installation, SDK authentication, or artifact storage configuration (use rhoai-mlflow), Training Hub, SDG Hub, Docling, or I

설치로 이동

소스 정보

저장소
adnan-drina/rhoai3-coding-demo
최근 소스 활동
2026년 7월 6일 18:02
감지된 SKILL.md 언어
영어
스타
3
포크
2

설치 방법

기본적으로 소스를 먼저 확인하는 Prompt가 선택됩니다. 직접 명령으로 전환하거나 로컬 사본을 다운로드할 수도 있습니다.

소스 파일 검토

설치 여부를 결정하기 전에 SKILL.md와 SkillsMP에 표시된 보조 파일을 읽어 보세요.