Skip to main content

module-4

This skill should be used when a learner is working through Module 4 ("Agent Customization") of the Build-an-Agent workshop and wants help understanding the concepts, the code, training, or GPU issues — e.g. "/module-4 should I train my agent or just prompt it?", "/module-4 what is GRPO?", "explain SFT vs GRPO", "how does the reward function work?", "what is reward hacking?", "help me with the GRPOConfig exercise", "my training crashes with OOM", "rewards aren't improving", "the reward server isn't responding", "what is NeMo Data Designer?", "how do I run the customized agent?". It turns the agent into a Module 4 learning assistant (tutor) that explains customization/RL concepts in the workshop's framing, gives graduated hints WITHOUT completing exercises or kicking off training runs, and troubleshoots SDG, the NeMo Gym reward server, GRPO/unsloth training, GPU/OOM, and the customized agent. Module 4 customizes a bash agent into a LangGraph CLI expert via synthetic data generation (NeMo Data Designer) + GRPO

Zur Installation springen

Quellinformationen

Repository
brevdev/workshop-build-an-agent
Letzte Quellaktivität
8. Juli 2026 um 04:04
Erkannte Sprache von SKILL.md
Englisch
Sterne
133
Forks
84

Installationsoptionen

Standardmäßig ist der Prompt ausgewählt, der zuerst die Quelle prüft. Sie können zu einem direkten Befehl wechseln oder eine lokale Kopie herunterladen.

Quelldateien prüfen

Lesen Sie SKILL.md und alle von SkillsMP angezeigten Begleitdateien, bevor Sie sich für eine Installation entscheiden.