Token Budget Planner
Set a monthly spend cap and get a per-day token budget.
- $25 Free
- 45 sec
- No signup
Set your monthly budget
Pick your default model
See your daily token allowance
You get: A per-day + per-session token budget that keeps you under your cap.
Your budget & habits
Cached input is billed at ~10% of standard input. Higher cache = more prompts per dollar.
What your budget buys
Prompts / mo
3.0k
Prompts / day
136
Sessions / mo
200
Hours / mo
300h
Tokens / mo
27.6M
Daily $ cap
$4.55
Cost per prompt at these settings
$0.033
Suggested budget split
Estimates using rounded per-million rates. Real spend depends on caching, tool use, and Anthropic's live pricing. Use this to set a target, then reconcile to your dashboard at month end.
Working backward from a budget
Most cost tools start with your usage and tell you what it will cost. This one runs the arrow the other way: you set a monthly dollar cap and it tells you how much Claude Code that actually buys - how many prompts, how many sessions, how many hours, at your chosen model. That is the more useful framing when you have a fixed budget and need to make it last. Instead of finding out on the 20th that you have blown through your cap, you start the month knowing your daily allowance and can pace to it deliberately. The planner divides your budget by the cost per prompt to get prompts per month, then converts that into the units you actually think in: sessions, hours, and a per-day spending line.
The one number everything hangs on
Cost per prompt is the hinge. It is computed from four things: your input tokens (billed at the model's input rate), your cached input (billed at roughly a tenth of that), your output tokens (billed at roughly 5x the input rate), and your model choice. Change any of them and the whole plan shifts. Halve your average prompt size and you double how many prompts your budget buys. Switch from Sonnet to Opus and you cut your prompt count by roughly 5x for the same money. This is why the planner exposes all four knobs - they are the levers that determine whether your budget feels generous or tight.
A budget is a pacing tool, not a ceiling to fear
The point of a per-day allowance is not to ration yourself into using Claude Code less. It is to spend deliberately - to know that a big Opus reasoning session costs several normal Sonnet prompts, and to make that trade on purpose rather than by accident.
How to read your allowance
- Prompts per month - the raw capacity your budget buys at your current settings.
- Prompts per day - the same figure spread across your working days. This is your daily pace target, the number to keep in your head.
- Sessions per month - prompts divided by how many prompts a typical session takes. Useful if you think in work sessions rather than individual prompts.
- Hours per month - sessions times hours per session. This translates the budget into calendar time, which is often the most intuitive unit of all.
- Daily dollar cap - your budget divided by working days. A simple guardrail: if you are past this by mid-afternoon, you are running hot.
The suggested budget split, and why
The allocation panel splits your budget across three tiers of work because using one model for everything is how budgets get wasted. The default split - the majority to Sonnet for core coding, a meaningful slice to Opus for genuinely hard reasoning, and a small slice to Haiku for bulk or cheap tasks - reflects how efficient users actually work. Sonnet is the daily driver that handles most coding at a fraction of Opus's cost. Opus is reserved for the moments where reasoning is the bottleneck and its price pays for itself in an avoided rewrite. Haiku mops up the high-volume, low-stakes work where speed and cheapness matter more than depth. Spending your whole budget on Opus is the single most common way to burn through it with nothing to show for the premium.
How to make your budget go further
- Default to Sonnet, escalate to Opus only when reasoning is genuinely the bottleneck. This alone can multiply how many prompts your budget buys.
- Raise your cache hit rate by referencing the same files in the same order across prompts. Cached input costs about a tenth of fresh input, so this is nearly free capacity.
- Cap your output. Output tokens cost roughly 5x input, so asking for a diff instead of a full-file rewrite is one of the highest-leverage habits you can build.
- Trim input. Point Claude Code at file paths and line ranges rather than pasting whole modules. Half of blown budgets are oversized inputs.
- Keep CLAUDE.md tight. It loads every session, so every line is a recurring cost against your budget.
- Compact long sessions. A bloated session reloads its whole history each turn, quietly eating your daily allowance.
Setting a realistic budget in the first place
If you do not have a number yet, work forward from value rather than fear. The ROI of Claude Code for professional work is usually strongly positive, so an artificially low budget often costs you more in lost output than it saves in spend. A sensible approach is to set a budget you are comfortable with, run a month, reconcile the estimate against your real dashboard usage, and adjust. The planner gives you the pacing; the dashboard gives you the truth. Together they let you tune a budget that is high enough to work well and low enough to stay disciplined.
What this planner does not model
This is a planning estimate. It assumes consistent prompt size and average cache behavior, and it uses rounded per-million rates. It does not model tool-call outputs that can spike a single prompt, MCP-loaded context, long-context pricing tiers, or the exact terms of a specific plan. The suggested split is a sensible default, not a prescription - your real allocation depends on your work. Use the planner to set a target and a daily pace, then reconcile to your actual usage at month end and adjust the inputs until the estimate tracks reality.
Frequently asked questions
How does the planner turn my budget into prompts?
It computes a cost per prompt from your model, prompt size, and cache rate, then divides your monthly budget by that figure. From prompts it derives sessions, hours, and a per-day allowance so you can pace to the number you find most intuitive.
Why does switching to Opus shrink my allowance so much?
Opus costs roughly 5x Sonnet per token, so the same budget buys about a fifth as many prompts. That is exactly why the suggested split reserves Opus for genuinely hard reasoning rather than everyday work.
What is the suggested budget split based on?
It reflects how efficient users actually work: most of the budget on Sonnet for core coding, a slice on Opus for hard reasoning where it pays for itself, and a small slice on Haiku for bulk cheap tasks. It is a sensible default, not a rule.
How do I raise how many prompts my budget buys?
Default to Sonnet over Opus, raise your cache hit rate, cap your output length, and trim oversized inputs. Each lowers your cost per prompt, which directly increases how many prompts the same budget affords.
Should I set a low budget to be safe?
Not reflexively. For professional work the ROI is usually strongly positive, so an artificially low cap can cost you more in lost output than it saves. Set a comfortable budget, run a month, reconcile to your dashboard, and adjust.
How accurate is the token and prompt count?
It is a planning estimate using rounded per-million rates and average cache assumptions. It does not model tool calls, MCP context, or long-context tiers. Use it to set a target, then reconcile against your real dashboard usage at month end.
Liked this tool? The club is the next step.
Join Claude Code Club for $9/month. 650+ lessons, weekly updates, and the workflows behind every tool on this site.
- No experience needed
- Cancel anytime
- Updated weekly
