Skip to main content
Run any Skill in Manus
with one click

llm-4bit-nf4-double-quantization

Stars52
Forks4
UpdatedApril 22, 2026 at 17:42

Load large LLMs with 4-bit NF4 quantization and optional double quantization via BitsAndBytes to reduce GPU memory by 4x while preserving inference quality

Installation

Install with Codex or Claude Copy this prompt, paste it into Codex, Claude, or another assistant, and let it review the skill page and install it for you.

SKILL.md
readonly