Skip to main content
Manusで任意のスキルを実行
ワンクリックで

fine-tuning-llms

スター23
フォーク7
更新日2026年5月7日 21:51

Deep LLM fine-tuning operational intuition — LoRA/QLoRA merge pitfalls, preference optimization (DPO/ORPO/SimPO/GRPO), chat-template footguns, completion-only loss masking, sequence packing, eval contamination, toolchain (TRL/axolotl/unsloth/torchtune). Load when post-training open-weights models, picking preference-optimization methods, configuring LoRA/QLoRA, suspecting eval contamination, or choosing a toolchain. Skip for inference/serving (use `llm-inference-serving`), prompt engineering, RAG, or pretraining. Triggers on: "QLoRA", "merge_and_unload", "DPO", "ORPO", "SimPO", "GRPO", "GRPOTrainer", "axolotl", "unsloth", "torchtune", "nf4", "apply_chat_template", "DataCollatorForCompletionOnlyLM", "cu_seqlens", "MMLU contamination", "catastrophic forgetting".

インストール

Codex または Claude でインストール この Prompt をコピーして Codex、Claude、または他のアシスタントに貼り付けると、Skill ページを確認してインストールできます。

SKILL.md
readonly