Skip to main content

rhoai-model-deployment

Use when documenting, reviewing, deploying, or operating Red Hat OpenShift AI model deployment user workflows from the official Deploying models guide: model storage in S3, URI, PVC, or OCI containers/modelcars; building OCI model images; dashboard Deploy a model wizard; generative and predictive model deployment; automatic serving-runtime selection; hardware profile matching; manual runtime selection; administrator llm-d override behavior; Rolling update versus Recreate deployment strategy; AI asset endpoint registration; external routes; token authentication; MLServer deployment settings; CLI deployment from OCI images with ServingRuntime and InferenceService; NVIDIA NIM model deployment; endpoint token lookup; and runtime-specific inference endpoint paths for Caikit, OpenVINO, vLLM, Triton, and MLServer. Do NOT use for model-serving platform enablement, runtime template configuration, vLLM runtime arguments, or NIM platform enablement (use rhoai-model-serving-platform), deployed model metrics and day-2 ope

الانتقال إلى التثبيت

معلومات المصدر

المستودع
adnan-drina/rhoai3-coding-demo
آخر نشاط في المصدر
٦ يوليو ٢٠٢٦ في ١٩:٢٨
لغة SKILL.md المكتشفة
الإنجليزية
النجوم
٣
التفرعات
٢

خيارات التثبيت

يُحدَّد Prompt الذي يراجع المصدر أولًا بشكل افتراضي. يمكنك التبديل إلى أمر مباشر أو تنزيل نسخة محلية.

مراجعة ملفات المصدر

اقرأ SKILL.md وأي ملفات مرافقة يعرضها SkillsMP قبل أن تقرر التثبيت.