Slime Rl Training

Provides guidance for LLM post-training with RL using slime, a Megatron+SGLang framework. Use when training GLM models, implementing custom data generation workflows, or needing tight Megatron-LM integration for RL scaling.

braxtonROSE4 4962faf 3 files · 29.8 KB Updated

File contents

braxtonROSE4/zorro-agent/tree/main/optional-skills/mlops/slime commit 4962faf38a

Frequently asked questions

npx skillmds@latest add braxtonrose4/slime-rl-training