Huggingface Trl Training

Train and fine-tune transformer language models using TRL (Transformers Reinforcement Learning). Supports SFT, DPO, GRPO, KTO, RLOO and Reward Model training via CLI commands.

pyfagorass d1b3ea4 8.3 KB Updated

File contents

pyfagorass/bookofspells/tree/main/skills/huggingface/trl-training commit d1b3ea48b8

Frequently asked questions

npx skillmds@latest add pyfagorass/huggingface-trl-training