Skip to main content

rhoai-model-evaluation

Guide RHOAI evaluation workflows once active evaluation content exists; during the reimplementation, use this skill to rebuild EvalHub, RAG evaluation, KFP/MLflow evidence, RAGAS patterns, and standard model benchmarking workflows from legacy references. Use when the user asks to run evaluation, evaluate a model, benchmark model performance, check RAG quality, compare pre-RAG vs post-RAG answers, run LM-Eval, create an LMEvalJob, interpret eval results, add new test questions, or modify the judge prompt. Also use when eval pipelines fail, LMEvalJob pods are stuck, or evaluation reports show unexpected scores. Do NOT use for official product EvalHub, LM-Eval, LMEvalJob, or automated risk assessment workflows from the Red Hat evaluation guide (use rhoai-evaluation), product AutoRAG dashboard optimization runs, leaderboard review, or generated notebooks (use rhoai-autorag), MLflow platform installation, SDK authentication, or artifact storage configuration (use rhoai-mlflow), Training Hub, SDG Hub, Docling, or I

Ir para a instalação

Informações da origem

Repositório
adnan-drina/rhoai3-coding-demo
Última atividade na origem
6 de julho de 2026 às 18:02
Idioma detectado do SKILL.md
inglês
Estrelas
3
Forks
2

Opções de instalação

Por padrão, está selecionado o prompt que primeiro revisa a origem. Você pode mudar para um comando direto ou baixar uma cópia local.

Revise os arquivos de origem

Leia o SKILL.md e os arquivos complementares exibidos pelo SkillsMP antes de decidir se vai instalar.