Hugging Face Model Trainer

This skill should be used when users want to train or fine-tune language models using TRL (Transformer Reinforcement Learning) on Hugging Face Jobs infrastructure. Covers SFT, DPO, GRPO and reward modeling training methods, plus GGUF conversion for...

jorcan Updated 0 repo stars

File contents

jorcan/automejora_agentes/tree/main/active_skills/hugging_face_model_trainer commit fe3144cfce

Frequently asked questions

npx skillmds@latest add jorcan/hugging-face-model-trainer