Prompt Caching

Caching strategies for LLM prompts including Anthropic prompt caching, response caching, and Cache-Augmented Generation (CAG) to cut latency and cost. USE WHEN you want to reduce LLM cost/latency by caching prompt prefixes or responses.

Sheshiyer Updated

File contents

Sheshiyer/skill-clusters/tree/main/skills/prompt-caching commit ffde1dc5a2

Frequently asked questions

npx skillmds@latest add sheshiyer/prompt-caching