Skip to main content
Run any Skill in Manus
with one click

local-llm-gguf

Stars1
Forks0
UpdatedJune 26, 2026 at 03:12

从 HuggingFace 下载 GGUF 模型并在本地运行(llama.cpp 或 Ollama)。 当用户问"怎么本地跑模型"、"部署 GGUF"、"llama.cpp 安装"、"本地跑 LLM"、 "下载 HuggingFace 模型"、"量化模型怎么选"时使用。 也适用于:检查本地 LLM 环境(ollama/llama.cpp 版本、安装路径、模型路径)、 迁移 Ollama 模型到其他磁盘、配置 OLLAMA_MODELS 环境变量。 覆盖:llama.cpp 安装(Windows)、GGUF 下载、量化版本选择、 llama-server/llama-cli 启动参数、Ollama 替代方案、常见问题排查。 也适用于:计算显存与上下文长度关系、估算模型显存需求、优化显存使用。 触发词:GGUF、llama.cpp、本地部署、本地运行、量化、Q4_K_M、Q8_0、gguf download、local LLM、 ollama 版本、ollama 路径、模型迁移、OLLAMA_MODELS、显存、上下文长度、VRAM、GPU内存、8GB显存。

Installation

Install with Codex or Claude Copy this prompt, paste it into Codex, Claude, or another assistant, and let it review the skill page and install it for you.

SKILL.md
readonly