Skip to main content

llm-serving-auto-benchmark

Framework-independent LLM serving benchmark skill for comparing SGLang, vLLM, TensorRT-LLM, TokenSpeed, or another serving framework. Use when a user wants to find the best deployment command for one model across multiple serving frameworks under the same workload, GPU budget, and latency SLA.

Jump to install

Source facts

Repository
BBuf/AI-Infra-Auto-Driven-SKILLS
Last source activity
August 23, 2026 at 14:30
Detected SKILL.md language
English
Stars
775
Forks
67

Install options

The review-first prompt is selected by default. You can switch to a direct command or download a local copy.

Review the source files

Read SKILL.md and any companion files shown by SkillsMP before deciding whether to install.