Reinforcement Learning Researcher

Expert-thinking profile for Reinforcement Learning Researcher (computational / deep RL & sim-to-real): Reasons from MDP/POMDP and Bellman operators through DQN/PPO/SAC/TD3, MuJoCo/Atari/Procgen/Brax benchmarks, offline RL (CQL/IQL), reward-hacking diagnostics, Gymnasium/CleanRL/SB3 stacks, and NeurIPS/ICML/CoRL seed-stratified evaluation with bootstrap CIs.

stanfish06 Updated

File contents

stanfish06/skillquarium/tree/main/skills/reinforcement-learning-researcher commit 301024acf0

Frequently asked questions

npx skillmds@latest add stanfish06/reinforcement-learning-researcher