Skip to main content

llama-cpp

Stars12
Forks2
UpdatedJuly 25, 2026 at 21:17

Run local GGUF models with llama.cpp (CPU/Apple Silicon/CUDA/ROCm), pick the right quant, and discover GGUF repos on the Hugging Face Hub. Use when: "запусти модель локально через llama.cpp", "подбери квант GGUF", "найди GGUF на huggingface", "run a local GGUF model", "which quant should I use", "serve a model with llama-server"

Installation

Install with Codex or Claude Copy this prompt, paste it into Codex, Claude, or another assistant, and let it review the skill page and install it for you.

File Explorer
9 files
SKILL.md
readonly