# Paleo Budget

> Use when user says "budget", "token limit", "stay under N tokens", or wants per-task token caps. Enforce a hard token budget per task/response — track estimate, stop before limit, summarize if exceeded. Off: "no budget" / "unlimited".

- Skill: `mocasus/paleo-budget` (Agent Skill)
- Install (CLI): `npx skillmds@latest add mocasus/paleo-budget`
- Raw SKILL.md: https://api.skillmd.com/api/skills/mocasus/paleo-budget/raw
- Safety review: pending
- Works with: Claude Code, Claude.ai, OpenAI Codex
- Category: Research & Search
- License: MIT
- Author: mocasus (https://skillmd.com/u/mocasus)
- Updated: 2026-09-17
- Page: https://skillmd.com/skills/mocasus/paleo-budget

---


# paleo-budget
Hard token budget per task. Cap spend, never overflow.

## Rules
- Set budget up front: "budget 2000" = max ~2000 out-tokens for the whole task.
- Estimate before writing: rough tok ≈ chars / 4. Track running count.
- Stop at ~90% budget. If answer incomplete, summarize remainder in bullets.
- Tight budget → pair with `paleo` (terse). budget + paleo = max save.
- Code/commands stay exact — budget hits prose, not correctness.

## Levels
- `soft`: warn near limit, keep going.
- `hard` (default): cut off at limit, summarize the tail.
- `strict`: fail-loud if estimate exceeds budget before starting.

## Switch
- `/budget 1500 soft|hard|strict` set cap + mode.
- "no budget" / "unlimited" → off.

## Gotchas
- Token estimate is approximate (chars/4). Leave 10% headroom.
- Never truncate the critical part of the answer — summarize instead.
- Budget = output cap. Input/context not counted unless user says "total".

