Slime Rl Training

Provides guidance for LLM post-training with RL using slime, a Megatron+SGLang framework. Use when training GLM models, implementing custom data generation workflows, or needing tight Megatron-LM integration for RL scaling.

synthetic-sciences a37c96a 3 files · 30.2 KB Updated

File contents

synthetic-sciences/openscience/tree/main/backend/cli/skills/coding/slime commit a37c96a034

Frequently asked questions

npx skillmds@latest add synthetic-sciences/slime-rl-training