Prompt Caching

Use prompt caching correctly across Anthropic, OpenAI, Bedrock, and Gemini to cut cost and latency on hot paths. Use when the user is building a production LLM app and mentions prompt caching, cache hits, cache key, cache TTL, ephemeral cache, system-prompt caching, or asks "why is my cache hit rate low?" / "should I cache this?".

cobusgreyling d051749 3 files · 10.2 KB Updated

File contents

cobusgreyling/agent-skills/tree/main/skills/prompt-caching commit d051749edb

Frequently asked questions

npx skillmds@latest add cobusgreyling/prompt-caching