Grpo Finetune

Fine-tune a model with GRPO on Fireworks-managed GPUs from a plain-English task description and a dataset. Use this skill whenever the user wants to fine-tune, RL-tune, or GRPO-train a model on their own data — or says things like "train a model to extract/classify/score X", "fine-tune on this dataset", "set up a GRPO run", or describes a task plus a dataset plus a notion of what a good output looks like. Trigger even when the user does not name GRPO or Fireworks explicitly.

gabrielmoreira Updated 17 repo stars

File contents

gabrielmoreira/agent-skills-mirror/tree/main/mirrors/repos/patchy631@ai-engineering-hub/grpo-finetuning-qwen3/agent-skill/grpo-finetune commit 2a8f6d242f

Frequently asked questions

npx skillmds@latest add gabrielmoreira/grpo-finetune