Skip to main content
Jeden Skill in Manus ausführen
mit einem Klick

llm-4bit-nf4-double-quantization

Sterne52
Forks4
Aktualisiert22. April 2026 um 17:42

Load large LLMs with 4-bit NF4 quantization and optional double quantization via BitsAndBytes to reduce GPU memory by 4x while preserving inference quality

Installation

Mit Codex oder Claude installieren Kopieren Sie diesen Prompt, fügen Sie ihn in Codex, Claude oder einen anderen Assistant ein und lassen Sie die Skill-Seite prüfen und installieren.

SKILL.md
readonly