Skip to main content

rhoai-distributed-inference-llmd

Use when documenting, reviewing, or rebuilding Red Hat OpenShift AI Distributed Inference with llm-d from the official deployment guide: LLMInferenceService resources, Gateway API discovery and selection, OpenShift Gateway Controller requirements, LeaderWorkerSet Operator prerequisites, Red Hat Connectivity Link and Kuadrant authentication, Authorino TLS setup, security.opendatahub.io/enable-auth, ServiceAccount JWT inference access, vLLM argument configuration for llm-d deployments, Endpoint Picker scheduler settings, ConfigMap-based scheduler references, Workload Variant Autoscaler enablement, VariantAutoscaling resources, saturation scaling ConfigMaps, mixed-workload flow control, InferenceObjective priority queuing, and llm-d observability. Do NOT use for standard single-model KServe deployment without llm-d (use rhoai-model-deployment), model-serving platform/runtime setup (use rhoai-model-serving-platform), MaaS governance over llm-d endpoints (use rhoai-maas-governance), Kueue/Ray/Training distributed

跳到安装

来源信息

仓库
adnan-drina/rhoai3-coding-demo
最近来源活动
2026年7月6日 18:02
检测到的 SKILL.md 语言
英语
星标
3
分支
2

安装方式

默认使用会先检查来源的 Prompt;你也可以切换为直接命令,或下载本地副本。

检查来源文件

决定是否安装前,请先阅读 SKILL.md,以及 SkillsMP 当前展示的配套文件。