Densegrpo Flow Matching

Improve diffusion model alignment by assigning step-wise rewards during denoising instead of terminal rewards. Fixes sparse reward signal mismatch in multi-step generation processes through ODE-based reward estimation.

adu2021 de1f811 7.2 KB Updated

File contents

adu2021/skillxiv/tree/main/skills/skillxiv-v0.0.2-claude-opus-4.6/densegrpo-flow-matching commit de1f811de8

Frequently asked questions

npx skillmds@latest add adu2021/densegrpo-flow-matching