Transformer Attention

Use when reasoning about Transformer self-attention, multi-head attention, positional encoding, masked decoder attention, or why attention replaced recurrence/convolutions in sequence models; not for generic NLP or unrelated attention topics.

vectifyai a5b9b92 2 files · 6.0 KB Updated

File contents

vectifyai/openkb/tree/main/examples/skills/transformer-attention commit a5b9b92385

Frequently asked questions

npx skillmds@latest add vectifyai/transformer-attention