Skip to main content

qad

Run explicitly requested ModelOpt Quantization-Aware Distillation (QAD) on Slurm through Megatron Bridge to recover a measured BF16-to-PTQ accuracy gap. Use only when the user explicitly asks for QAD, including its topology, data preparation, Slurm launch, resume, checkpoint export, or recovery decisions.

Jump to install

Source facts

Repository
NVIDIA/Model-Optimizer
Last source activity
August 12, 2026 at 16:33
Detected SKILL.md language
English
Stars
3,604
Forks
567

Install options

The review-first prompt is selected by default. You can switch to a direct command or download a local copy.

Review the source files

Read SKILL.md and any companion files shown by SkillsMP before deciding whether to install.