Agentflow In The Flow Optimization

Decompose agent work across four specialized modules (planner, executor, verifier, generator) coordinated via evolving memory. Use Flow-GRPO to convert multi-turn sparse-reward optimization into sequential single-turn updates with outcome broadcasting, achieving 4-15% accuracy gains on benchmarks while scaling better than monolithic agent policies.

adu2021 79b428e 10.3 KB Updated

File contents

adu2021/skillxiv/tree/main/skills/skillxiv-v0.0.2-claude-opus-4.6/agentflow-in-the-flow-optimization commit 79b428e5ea

Frequently asked questions

npx skillmds@latest add adu2021/agentflow-in-the-flow-optimization