Slime Rl Training

Provides guidance for LLM post-training with RL using slime, a Megatron+SGLang framework. Use when training GLM models, implementing custom data generation workflows, or needing tight Megatron-LM integration for RL scaling.

tianhao909 Updated 1 repo stars

File contents

tianhao909/AI-Research-SKILLs-cn/tree/main/06-post-training/slime commit 8f5a17b8fb

Frequently asked questions

npx skillmds@latest add tianhao909/slime-rl-training