Train Sft

Supervised fine-tune a causal LM with TRL `SFTTrainer`. Triggered when the user wants to fine-tune / SFT / instruct-tune / chat-tune a model on conversational, prompt-completion, or text-formatted data. Enforces the literature-first → audit → smoke-test → scale workflow.

mybigday bc1721c 2.6 KB Updated

File contents

mybigday/ml-intern-kit/tree/main/.claude/skills/train-sft commit bc1721c04f

Frequently asked questions

npx skillmds@latest add mybigday/train-sft