원클릭으로
model-quantization
""Quantize and optimize ML models for edge deployment across all four tiers. Use when: converting models to LiteRT, ONNX export, INT8/FP16 quantization, GGUF conversion, MLX format, BitNet 1-bit quantization, model size reduction, inference optimization, Adreno GPU delegate, NNAPI configuration, profiling model performance on target hardware.""
Codex 또는 Claude로 설치 이 Prompt를 복사해 Codex, Claude 또는 다른 어시스턴트에 붙여 넣으면 Skill 페이지를 검토하고 설치를 진행할 수 있습니다.