Dynamic Reward Scaling and Normalization

Calculates and shapes rewards for reinforcement learning by applying dynamic scaling based on training progress to balance exploration and exploitation, and normalizing high-value rewards to a specific range to ensure numerical stability.

ECNU-ICALK Updated 559 repo stars

File contents

ECNU-ICALK/AutoSkill/tree/main/SkillBank/ConvSkill/english_gpt4_8/dynamic-reward-scaling-and-normalization commit 89ffc9ea15

Frequently asked questions

npx skillmds@latest add ecnu-icalk/dynamic-reward-scaling-and-normalization