Sparse Autoencoder Training

Fornece orientação para treinar e analisar Autoencodificadores Esparsos (SAEs) usando SAELens para decompor ativações de redes neurais em features interpretáveis. Use ao descobrir features interpretáveis, analisar superposição ou estudar representações monossemânticas em modelos de linguagem.

artubss Updated 10 repo stars

File contents

artubss/SKILLS-CLAUDE-CODE/tree/main/skills/skills/ai-research/mechanistic-interpretability-saelens commit 3c01def1a1

Frequently asked questions

npx skillmds@latest add artubss/sparse-autoencoder-training