Long Context Vs RAG

Decision framework: when long-context (Gemini 2M, Claude 200k, GPT 128k) beats RAG, hybrid approaches (RAG narrows, long-context reads), cost-quality-latency tradeoffs, lost-in-the-middle / context rot research, needle vs synthesis tasks, prompt caching economics, concrete $ per query math. USE WHEN: user mentions "long context vs RAG", "Gemini 2M", "lost in the middle", "context rot", "when not to use RAG", "stuff the prompt", "prompt caching cost" DO NOT USE FOR: implementing RAG - use `rag-architecture`; evaluating RAG - use `rag-evaluation`; chunking decisions alone - use `chunking-strategies`

claude-dev-suite Updated 28 repo stars

File contents

claude-dev-suite/claude-dev-suite/tree/main/skills/rag/long-context-vs-rag commit b58feeffea

Frequently asked questions

npx skillmds@latest add claude-dev-suite/long-context-vs-rag