LLM Response Caching Layer

Implement semantic and exact-match caching for LLM responses to reduce cost 40-60% and latency. Activate on: LLM caching, semantic cache, reduce API costs, cache AI responses. NOT for: general web caching (caching-strategies), CDN config (cloudflare-worker-dev).

curiositech Updated 10 repo stars

File contents

curiositech/windags-skills/tree/main/skills/llm-response-caching-layer commit 32843f1478

Frequently asked questions

npx skillmds@latest add curiositech/llm-response-caching-layer