Skip to main content
Execute qualquer Skill no Manus
com um clique

alchemist-playbook

Estrelas5
Forks0
Atualizado11 de julho de 2026 às 11:19

Evidence-based training-recipe advisor (煉丹調參) distilled from published runs: LLaMA 1-3, OLMo 1-3, DeepSeek-V3, SmolLM2, Kimi K2, GLM-5, PivotRL, Agents-A1, EvoLM, LFM2, VibeThinker, FAC-Synthesis, Zephyr, Tulu 3, SimPO, ORPO, QLoRA, Whisper, OWSM, wav2vec 2.0, HuBERT. Use whenever the user asks about training hyperparameters (learning rate, batch size, warmup, scheduler, optimizer, beta, weight decay, epochs), debugging a training run (loss spike, NaN, divergence, slow convergence, overfitting), designing a pretraining/SFT/DPO/RLHF/RLVR/agentic-RL/LoRA/QLoRA/speech (ASR/TTS) recipe, compute or token budgets, distillation/curriculum/model-merging strategy, small or on-device models, what to monitor or which benchmarks/eval suite per stage, capability regression or catastrophic forgetting, tracking/reporting training runs (實驗追蹤, HTML run reports), or mentions 煉丹, 調參, "training recipe", "fine-tuning settings", "what LR should I use", "eval suite", "release gate" — even if they never say "hyperparameter".

Instalação

Instalar com Codex ou Claude Copie este prompt, cole no Codex, Claude ou outro assistente e deixe que ele revise a página da skill e instale para você.

Explorador de arquivos
20 arquivos
SKILL.md
readonly