一键导入
model-quantization
""Quantize and optimize ML models for edge deployment across all four tiers. Use when: converting models to LiteRT, ONNX export, INT8/FP16 quantization, GGUF conversion, MLX format, BitNet 1-bit quantization, model size reduction, inference optimization, Adreno GPU delegate, NNAPI configuration, profiling model performance on target hardware.""
用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。