Inputs And Layer Wise States

Analyze and visualize layer-wise gradient behaviors in LLMs during fine-tuning for fast vs slow thinking tasks, calculate gradient statistics, and understand training patterns across different model layers

zjunlp aa5d9d7 4 files · 45.4 KB Updated

File contents

zjunlp/mechanist/tree/main/skills/mechanism-skills/gradient-detection/inputs-and-layer-wise-states commit aa5d9d7856

Frequently asked questions

npx skillmds@latest add zjunlp/inputs-and-layer-wise-states