Transformer Lens Interpretability

Provides guidance for mechanistic interpretability research using TransformerLens to inspect and manipulate transformer internals via HookPoints and activation caching. Use when reverse-engineering model algorithms, studying attention patterns, or performing activation patching experiments.

synthetic-sciences 2477715 4 files · 31.2 KB Updated

File contents

synthetic-sciences/openscience/tree/main/backend/cli/skills/ml-training/transformer-lens commit 2477715ae5

Frequently asked questions

npx skillmds@latest add synthetic-sciences/transformer-lens-interpretability