Pufferlib

High-performance reinforcement learning framework optimized for speed and scale. Use when you need fast parallel training, vectorized environments, multi-agent systems, or integration with game environments (Atari, Procgen, NetHack). Achieves 2-10x speedups over standard implementations. For quick prototyping or standard algorithm implementations with extensive documentation, use stable-baselines3 instead.

mkurman a1b2024 8 files · 97.2 KB Updated

File contents

mkurman/zorai/tree/main/skills/scientific-skills/pufferlib commit a1b2024ae3

Frequently asked questions

npx skillmds@latest add mkurman/pufferlib