Rwml Reinforcement World Models

Train LLM agents to anticipate environment consequences by learning world models through reinforcement learning with embedding-space similarity rewards, avoiding task-specific labels while enabling robust environment adaptation.

adu2021 4ddc4b7 7.7 KB Updated

File contents

adu2021/skillxiv/tree/main/skills/skillxiv-v0.0.2-claude-opus-4.6/rwml-reinforcement-world-models commit 4ddc4b76fb

Frequently asked questions

npx skillmds@latest add adu2021/rwml-reinforcement-world-models