kernel-microbenchmark
Build, debug, and interpret vLLM GPU kernel microbenchmarks for CUDA, Triton, and CuteDSL, including CUPTI timing, correctness checks, generated-code inspection, multi-GPU measurements, and SOL sanity checks.
Source facts
- Repository
- vllm-project/vllm
- Last source activity
- August 25, 2026 at 07:25
- Detected SKILL.md language
- English
- Stars
- 90,488
- Forks
- 21,428
Install options
The review-first prompt is selected by default. You can switch to a direct command or download a local copy.
Review the source files
Read SKILL.md and any companion files shown by SkillsMP before deciding whether to install.