Trl Training

Post-train LLMs with TRL (Transformers Reinforcement Learning) — SFT, DPO, GRPO, KTO, and reward-model training. Use when writing or debugging training code with the TRL Python API or the trl CLI.

Hugging Face Updated 10.8k repo stars

File contents

huggingface/trl/tree/main/skills/trl-training commit 9046f03f97

Frequently asked questions

npx skillmds@latest add huggingface/trl-training-3