Token Diet

Production-ready token optimization: reduce costs 40–75% through retrieval pruning, smart caching, and model routing. Use whenever optimizing API costs, latency, or managing long context—especially for RAG pipelines, high-volume systems, multi-turn conversations, or when context exceeds 2K tokens.

VDADev2022 ab2df9c 4 files · 22.6 KB Updated

File contents

VDADev2022/token-diet/tree/main/ commit ab2df9c642

Frequently asked questions

npx skillmds@latest add vdadev2022/token-diet