Skip to main content

onnx-export-quantization

Stars13
Forks1
UpdatedJune 2, 2026 at 20:12

Use this skill when exporting ONNX models with mobius and quantizing them with Olive for deployment. Covers the mobius CLI, EP options, INT4 quantization (Q4_K_M and NF4), HuggingFace upload structure, GPU-accelerated quantization, common issues, and testing quantized models.

Installation

Install with Codex or Claude Copy this prompt, paste it into Codex, Claude, or another assistant, and let it review the skill page and install it for you.

SKILL.md
readonly