Sae Feature Activation State

Use this skill when working with Sparse Autoencoders (SAEs) for feature analysis, particularly for studying feature splitting, absorption, and attribution in language models. Activate for tasks involving SAE feature ablation, probing experiments, or analyzing how SAE latents affect model outputs in spelling and token-level tasks.

zjunlp d5e7a65 5 files · 77.9 KB Updated

File contents

zjunlp/mechanist/tree/main/skills/mechanism-skills/probing/sae-feature-activation-state commit d5e7a65ce3

Frequently asked questions

npx skillmds@latest add zjunlp/sae-feature-activation-state