Kv Cache Steering Reasoning

Guide frozen language models toward multi-step reasoning by modifying cached key-value representations after the prefilling stage. Extract steering vectors from contrastive prompt pairs and apply them to KV cache with scalar coefficients. Improves reasoning on GSM8K, ARC, CommonsenseQA while adding only 10ms overhead per token.

adu2021 e96b1ab 15.7 KB Updated

File contents

adu2021/skillxiv/tree/main/skills/skillxiv-v0.0.2-claude-opus-4.6/kv-cache-steering-reasoning commit e96b1abefb

Frequently asked questions

npx skillmds@latest add adu2021/kv-cache-steering-reasoning