Skip to main content

rhoai-distributed-inference-llmd

Use when documenting, reviewing, or rebuilding Red Hat OpenShift AI Distributed Inference with llm-d from the official deployment guide: LLMInferenceService resources, Gateway API discovery and selection, OpenShift Gateway Controller requirements, LeaderWorkerSet Operator prerequisites, Red Hat Connectivity Link and Kuadrant authentication, Authorino TLS setup, security.opendatahub.io/enable-auth, ServiceAccount JWT inference access, vLLM argument configuration for llm-d deployments, Endpoint Picker scheduler settings, ConfigMap-based scheduler references, Workload Variant Autoscaler enablement, VariantAutoscaling resources, saturation scaling ConfigMaps, mixed-workload flow control, InferenceObjective priority queuing, and llm-d observability. Do NOT use for standard single-model KServe deployment without llm-d (use rhoai-model-deployment), model-serving platform/runtime setup (use rhoai-model-serving-platform), MaaS governance over llm-d endpoints (use rhoai-maas-governance), Kueue/Ray/Training distributed

インストールへ移動

ソース情報

リポジトリ
adnan-drina/rhoai3-coding-demo
ソースの最終更新活動
2026年7月6日 18:02
検出された SKILL.md の言語
英語
スター
3
フォーク
2

インストール方法

デフォルトでは、最初にソースを確認する Prompt が選択されています。直接コマンドに切り替えるか、ローカルコピーをダウンロードすることもできます。

ソースファイルを確認

インストールを決める前に、SKILL.md と SkillsMP に表示されている付属ファイルをお読みください。