Skip to main content

llm-d-workload-tuner

This is an experimental Skill. It automatically tunes GKE vLLM inference server parameters and resources based on workload profiles specified in the benchmark configs.

Jump to install

Source facts

Repository
GoogleCloudPlatform/accelerated-platforms
Last source activity
August 20, 2026 at 16:47
Detected SKILL.md language
English
Stars
102
Forks
36

Install options

The review-first prompt is selected by default. You can switch to a direct command or download a local copy.

Review the source files

Read SKILL.md and any companion files shown by SkillsMP before deciding whether to install.