LLM Cost Latency Budget

Model the cost and latency of an LLM feature before it ships and surprises the bill. Use when asked to estimate LLM API costs, set a latency/token budget, decide which model tier to use, or bring down the cost of an AI feature. Produces a cost & latency budget — token math per request, monthly cost projection, model tiering, caching/streaming levers, p95 latency targets, and a guardrail/alert plan.

gabrielmoreira Updated 17 repo stars

File contents

gabrielmoreira/agent-skills-mirror/tree/main/mirrors/repos/mohitagw15856@pm-claude-skills/skills/llm-cost-latency-budget commit 05980d0967

Frequently asked questions

npx skillmds@latest add gabrielmoreira/llm-cost-latency-budget