Grpo Finetune

> Fine-tune a model with GRPO on Fireworks-managed GPUs from a plain-English task description and a dataset. Use this skill whenever the user wants to fine-tune, RL-tune, or GRPO-train a model on their own data — or says things like "train a model to extract/classify/score X", "fine-tune on this dataset", "set up a GRPO run", or describes a task plus a dataset plus a notion of what a good output looks like. Trigger even when the user does not name GRPO or Fireworks explicitly.

patchy631 Updated

File contents

patchy631/ai-engineering-hub commit c174478cc5

Frequently asked questions

npx skillmds@latest add patchy631/grpo-finetune