Set a spending limit for every customer, or team, or agent. Stop overspend the moment it happens, not when the invoice lands. And see exactly what each one cost, so you can bill it back.
The provider sends one bill. Your customers are many. Everything you can't see or stop lives in that gap.
The provider bill is a single number. You've no way to know what any one customer, team, or feature actually cost you.
A single customer on a flat plan can quietly burn the profit you made on ten others, and you find out when the invoice arrives.
Your pricing says "500 AI actions a month." Nothing actually stops someone using 5,000 and handing you the bill.
No rebuild. Point your existing AI calls through us, and set limits on the customers you already have.
Map a budget to each customer, team, or agent. Free gets £5, Pro gets £50, Enterprise whatever you choose — tied to the plans you already sell.
We check every call against the budget and block it the instant the limit's reached — before the cost lands, not after the invoice.
Get the exact cost of every customer, ready to drop into an invoice. The number your provider's bill never breaks down for you.
Works with your existing OpenAI or Anthropic setup — one line of config, or our API. See the docs →
Most teams are well served by soft caps. Turn on hard caps for the budgets where going over isn't an option.
Spend is tracked as it happens, and calls are blocked the moment a budget is hit. Works with every model and workload.
For customer-billed AI, regulated spend, or anything high-stakes. We reserve each call's worst-case cost before it runs — like a hotel holding a deposit on your card at check-in — then settle to the real amount after. Concurrent calls can't collectively bust the limit.
Set up in an afternoon. Works with your existing providers.
A budget for every customer, team, or agent — enforced automatically, in real time.
Calls are blocked the moment a limit is hit. Decisions land in milliseconds, not minutes.
See exactly what each customer used. Export it, bill on it, prove it.
Block at the limit, or reserve cost up front for a guarantee that holds under load — per budget.
OpenAI, Anthropic, Gemini, Mistral, Cohere, DeepSeek — one integration, pricing kept current.
Slack, webhook, or email when a budget is hit or crosses 80%.
Velocity limits catch runaway loops before a bug burns a whole budget.
Every allowed and blocked call, durably recorded — the record behind what you bill.
We record usage and cost, never the contents of your calls.
A real billing report — daily and monthly breakdowns, cost per call, CSV export. Not a mockup.
Every row is durably recorded with timestamps. The same data you bill on is the data you defend a chargeback with — one record your finance team can audit on demand.
Simple monthly tiers. Hard caps available as an add-on whenever you need them.
The things people ask before trying it. The full detail lives in the docs.
Set a budget for every customer, enforce it in real time, and bill on exactly what was used.