Pytorch Fsdp2

Adds PyTorch FSDP2 (fully_shard) to training scripts with correct init, sharding, mixed precision/offload config, and distributed checkpointing. Use when models exceed single-GPU memory or when you need DTensor-based sharding with DeviceMesh. Use when this capability is needed.

tomevault-io 0bb838f 2 files · 11.4 KB Updated

File contents

tomevault-io/skills-registry/tree/main/orchestra-research--ai-research-skills--pytorch-fsdp2 commit 0bb838f290

Frequently asked questions

npx skillmds@latest add tomevault-io/pytorch-fsdp2