Saelens

Use this skill when working with Sparse Autoencoders (SAEs) for mechanistic interpretability of language models, including training SAEs, loading pre-trained SAEs, analyzing neural network features, or integrating SAEs with TransformerLens, HuggingFace Transformers, or other PyTorch-based models.

zjunlp 64c2347 2 files · 3.1 KB Updated

File contents

zjunlp/mechanist/tree/main/skills/mechanism-skills/feature-dictionary-learning/SAE commit 64c2347cd7

Frequently asked questions

npx skillmds@latest add zjunlp/saelens