Finetuning Technique

Selects a fine-tuning technique (SFT, DPO, RLVR, or RLAIF) for the user's use case and validates it against the selected model's available recipes. Use when the user has decided to finetune and needs to choose a technique, or when the technique needs to be validated against a model. Requires a base model to already be selected (via model-selection skill).

gabrielmoreira Updated 17 repo stars

File contents

gabrielmoreira/agent-skills-mirror/tree/main/mirrors/repos/awslabs@agent-plugins/plugins/sagemaker-ai/skills/finetuning-technique commit 1ccb324b10

Frequently asked questions

npx skillmds@latest add gabrielmoreira/finetuning-technique