LLM spend monitoring for AI teams
TokenBurn watches your OpenAI and Anthropic spend continuously and tells you, in plain language, what changed, why, and who should care. Spot the ember before it becomes the fire.
5-minute setup · Admin key used only for read-only endpoints · No credit card

$1,284
+18% vs yesterday
claude-opus-5 · prod-assistant
+41% in 3 days
The problem
A runaway agent loop. A prompt that never hits the cache. An eval job pointed at your most expensive model. A forgotten API key still burning in someone's side project. Nobody notices until the invoice arrives — and by then, you're arguing about a surprise bill instead of shipping.
Most usage pages are dashboards. Dashboards need to be remembered. Engineers don't have time to remember. TokenBurn inverts this — it watches for you, and shows up in your morning routine only when something actually changed.
How it works
No SDK to wire in. No proxy in front of your requests. Just an Admin API key that we only use against the read-only usage and cost endpoints.
Email or Google. One click, nothing to install on your side.
Create an Admin API key in OpenAI or Anthropic and paste it once. It's encrypted at rest and only ever used for read-only usage and cost endpoints.
Watch every project and workspace in the org, or only the ones you care about. Change it any time.
MTD spend, projected month-end, tokens in and out, top movers, and the embers we've spotted — all before your first coffee.
Features
We surface the top movers in the last 7 days with plain-English explanations. "claude-opus-5 in prod-assistant is up 41% — cache hit rate fell off a cliff on Tuesday."
See every organization broken down by project or workspace, model, and line item — input, output, cached input, cache write. Drill into any spike in two clicks.
Know who is burning. Tokens and estimated cost per user and per API key, so a runaway loop has a name attached before lunch.
Enter your monthly budget and we show days of runway alongside MTD burn. Founders think in runway — we do too.
One email every morning with exactly what changed and why. No dashboards to remember.
Admin keys are encrypted at rest and never shown again — not even to you. We only call read-only usage and cost endpoints, and you can delete a key any time.
The metaphor
Founders track burn because it decides how long they survive. AI teams should treat token spend the same way — as a single number that changes every hour, and that someone must actually watch.
TokenBurn is that someone.
Cache discount left on the table
up to 90%
Cached input vs. full price
Time to detect a runaway loop
Days
Checking usage pages by hand
Time to detect with TokenBurn
< 24h
Daily digest + ember alerts
Setup time
~5 min
Sign in + paste an Admin key
Illustrative figures — your numbers depend on your traffic and models.
Pricing
No percent-of-savings fees, ever. Your savings should be your savings.
Free
in beta
$99
/ month (later)
Talk to us
volume + SSO
Sign in, paste an Admin API key, and wake up to clarity.