Slime User

Guide for using SLIME (LLM post-training framework for RL Scaling). Use when working with SLIME for reinforcement learning training of language models, including setup, configuration, training execution, multi-turn interactions, custom reward models, tool calling scenarios, or troubleshooting SLIME workflows. Covers GRPO, GSPO, PPO, Reinforce++, multi-agent RL, VLM training, FSDP/Megatron backends, SGLang integration, dynamic sampling, and custom generation functions.

tools-only bd85d6b 3 files · 26.1 KB Updated 7 repo stars

File contents

tools-only/X-Skills/tree/main/development/backend/979-skill_8395d2fc commit bd85d6b15e

Frequently asked questions

npx skillmds add tools-only/slime-user