Attention Head Output

Use this skill when you need to analyze and interpret the internal workings of large language models layer by layer, visualize hidden states and predictions across transformer layers, or understand how models like Llama-3.1-8B and Qwen-2.5-7B make predictions at each layer using the Logit Lens technique

zjunlp a8a3ed5 4 files · 62.2 KB Updated

File contents

zjunlp/mechanist/tree/main/skills/mechanism-skills/vocabulary-projection/attention-head-output commit a8a3ed5f3d

Frequently asked questions

npx skillmds@latest add zjunlp/attention-head-output