Module 4

This skill should be used when a learner is working through Module 4 ("Agent Customization") of the Build-an-Agent workshop and wants help understanding the concepts, the code, training, or GPU issues — e.g. "/module-4 should I train my agent or just prompt it?", "/module-4 what is GRPO?", "explain SFT vs GRPO", "how does the reward function work?", "what is reward hacking?", "help me with the GRPOConfig exercise", "my training crashes with OOM", "rewards aren't improving", "the reward server isn't responding", "what is NeMo Data Designer?", "how do I run the customized agent?". It turns the agent into a Module 4 learning assistant (tutor) that explains customization/RL concepts in the workshop's framing, gives graduated hints WITHOUT completing exercises or kicking off training runs, and troubleshoots SDG, the NeMo Gym reward server, GRPO/unsloth training, GPU/OOM, and the customized agent. Module 4 customizes a bash agent into a LangGraph CLI expert via synthetic data generation (NeMo Data Designer) + GRPO

NVIDIA 74e0533 7 files · 40.9 KB Updated 2.2k repo stars

File contents

nvidia/nemoclaw-community/tree/main/examples/recipes/nvidia/agentic-ai-learning-path/skills/module-4 commit 74e0533d5c

Frequently asked questions

npx skillmds@latest add nvidia/module-4