Slime Rl Training

Provides guidance for LLM post-training with RL using slime, a Megatron+SGLang framework. Use when training GLM models, implementing custom data generation workflows, or needing tight Megatron-LM integration for RL scaling.

peteedoo Updated 0 repo stars

File contents

peteedoo/project-template/tree/main/.agents/skills/imported/nousresearch--hermes-agent/slime commit 62fdc5b198

Frequently asked questions

npx skillmds@latest add peteedoo/slime-rl-training