Trl Training

Train and fine-tune transformer language models using TRL (Transformers Reinforcement Learning). Supports SFT, DPO, GRPO, KTO, RLOO and Reward Model training via CLI commands.

FrancoStino Updated 63 repo stars

File contents

FrancoStino/opencode-skills-collection/tree/main/bundled-skills/trl-training commit 017c3d4724

Frequently asked questions

npx skillmds@latest add francostino/trl-training