quark-torch-llm-eval
End-to-end LLM accuracy evaluation on AMD ROCm (ROCm-only) — docker container setup OR host (no-docker) runtime, vLLM/SGLang/ATOM serving, lm-eval / lighteval / evalscope benchmarks. Use when the user wants to evaluate, benchmark, or compare an LLM's accuracy. Trigger for "evaluate this model", "run gsm8k/mmlu/mmlu_pro/aime/gpqa/hellaswag/arc", "test accuracy", "measure perplexity", "compare quantized model accuracy", "does this mxfp4 model lose accuracy". For evaluating Quark Agent Skills themselves, use quark-torch-eval-runner instead.
Source facts
- Repository
- amd/Quark
- Last source activity
- July 9, 2026 at 06:10
- Detected SKILL.md language
- English
- Stars
- 166
- Forks
- 30
Install options
The review-first prompt is selected by default. You can switch to a direct command or download a local copy.
Review the source files
Read SKILL.md and any companion files shown by SkillsMP before deciding whether to install.