Skip to main content

rhoai-model-serving-platform

Use when documenting, reviewing, or configuring the Red Hat OpenShift AI model-serving platform from the official configuration guide: KServe-based model serving, NVIDIA NIM platform enablement, ServingRuntime and InferenceService concepts, supported and tested model-serving runtime posture, accelerator-specific runtimes, dashboard enablement of the model serving platform and runtimes, custom and tested runtime addition, speculative decoding, multi-modal vLLM configuration, runtime argument and environment-variable customization, vLLM KV cache configuration, and default deployment strategy. Do NOT use for deployed-model monitoring, KServe timeout tuning, multi-node vLLM operations, Kueue routing, Grafana dashboards, or NIM metrics operations (use rhoai-model-management-monitoring), installing RHOAI or KServe components (use rhoai-self-managed-installation), GPU enablement (use rhoai-nvidia-gpu-accelerators), project connection API annotations (use rhoai-project-workflows), project-scoped runtime templates (us

الانتقال إلى التثبيت

معلومات المصدر

المستودع
adnan-drina/rhoai3-coding-demo
آخر نشاط في المصدر
٦ يوليو ٢٠٢٦ في ١٨:٠٢
لغة SKILL.md المكتشفة
الإنجليزية
النجوم
٣
التفرعات
٢

خيارات التثبيت

يُحدَّد Prompt الذي يراجع المصدر أولًا بشكل افتراضي. يمكنك التبديل إلى أمر مباشر أو تنزيل نسخة محلية.

مراجعة ملفات المصدر

اقرأ SKILL.md وأي ملفات مرافقة يعرضها SkillsMP قبل أن تقرر التثبيت.