← all publishers

VDADev2022

@vdadev2022 source repo

1 published skill

  1. Token Diet · vdadev2022 bundle
    Production-ready token optimization: reduce costs 40–75% through retrieval pruning, smart caching, and model routing. Use whenever optimizing API costs, latency, or managing long context—especially for RAG pipelines, high-volume systems, multi-turn conversations, or when context exceeds 2K tokens.
    0
    installs