Slime Rl Training

Provides guidance for LLM post-training with RL using slime, a Megatron+SGLang framework. Use when training GLM models, implementing custom data generation workflows, or needing tight Megatron-LM integration for RL scaling.

photonics-dhl Updated

File contents

photonics-dhl/scholar-s-tea/tree/main/hermes-home/hermes-agent/optional-skills/mlops/slime commit 194910c9c8

Frequently asked questions

npx skillmds@latest add photonics-dhl/slime-rl-training