Sparse Autoencoder Training

Provides guidance for training and analyzing Sparse Autoencoders (SAEs) using SAELens to decompose neural network activations into interpretable features. Use when discovering interpretable features, analyzing superposition, or studying monosemantic representations in language models.

photonics-dhl 9e8eae6 4 files · 31.0 KB Updated

File contents

photonics-dhl/scholar-s-tea/tree/main/hermes-home/hermes-agent/optional-skills/mlops/saelens commit 9e8eae6037

Frequently asked questions

npx skillmds@latest add photonics-dhl/sparse-autoencoder-training