Skip to main content
在 Manus 中运行任何 Skill
一键导入

moe-finetune-limited-gpu

星标6
分支0
更新时间2026年7月11日 17:57

Run BEFORE fine-tuning ANY Mixture-of-Experts model (Qwen3-30B-A3B / -A22B, OLMoE, DeepSeek-MoE, any `*-A\d+B` model) on this fleet, and BEFORE dispatching a big-model (≥30B) paid RunPod training lane. Fires on "fine-tune a MoE", "MoE LoRA", "train Qwen3-30B-A3B", "expert routing / target the experts", "--n-cpu-moe / --cpu-moe", "which GPU for a 30B bf16 LoRA", "the RunPod train pod OOMed / disk-quota / got reaped / timed out", or "moe-sft-runpod / spark-moe-sft / mac-mlx-moe-train lane". This repo learned the MoE recipe AND the 30B-paid-lane infra recipe the expensive way (5 paid pods, ~$30, four one-per-pod infra walls) — front-load ALL of it so you do not rediscover it a paid pod at a time. Use even for a "quick" run and even if the 7B lane "already works" — MoE and 30B both break the small-dense defaults. Slash: /moe-finetune-limited-gpu. Augments wisdom-gpu-prebaked (RunPod deps) and spark-cluster-ops (pod cost-guard §5).

安装

用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。

SKILL.md
readonly