Skip to main content

rhoai-model-evaluation

Guide RHOAI evaluation workflows once active evaluation content exists; during the reimplementation, use this skill to rebuild EvalHub, RAG evaluation, KFP/MLflow evidence, RAGAS patterns, and standard model benchmarking workflows from legacy references. Use when the user asks to run evaluation, evaluate a model, benchmark model performance, check RAG quality, compare pre-RAG vs post-RAG answers, run LM-Eval, create an LMEvalJob, interpret eval results, add new test questions, or modify the judge prompt. Also use when eval pipelines fail, LMEvalJob pods are stuck, or evaluation reports show unexpected scores. Do NOT use for official product EvalHub, LM-Eval, LMEvalJob, or automated risk assessment workflows from the Red Hat evaluation guide (use rhoai-evaluation), product AutoRAG dashboard optimization runs, leaderboard review, or generated notebooks (use rhoai-autorag), MLflow platform installation, SDK authentication, or artifact storage configuration (use rhoai-mlflow), Training Hub, SDG Hub, Docling, or I

الانتقال إلى التثبيت

معلومات المصدر

المستودع
adnan-drina/rhoai3-coding-demo
آخر نشاط في المصدر
٦ يوليو ٢٠٢٦ في ١٨:٠٢
لغة SKILL.md المكتشفة
الإنجليزية
النجوم
٣
التفرعات
٢

خيارات التثبيت

يُحدَّد Prompt الذي يراجع المصدر أولًا بشكل افتراضي. يمكنك التبديل إلى أمر مباشر أو تنزيل نسخة محلية.

مراجعة ملفات المصدر

اقرأ SKILL.md وأي ملفات مرافقة يعرضها SkillsMP قبل أن تقرر التثبيت.