Book a free cost audit.
Thirty minutes on a call, then a written audit of where your AI spend is actually going. No cost, no commitment, and no software to install. A person does the work and a person presents it back.
1. What happens
- The call (30 minutes). You describe the stack: which providers, roughly what you're spending, and what triggered the question. We tell you on the call whether there is anything worth finding. If there isn't, we say so and you've lost half an hour.
- The audit (about a week). We take read-only billing and usage exports and map the spend to the workloads producing it. This is the part almost nobody has done, because the invoice arrives as one line per provider.
- The readout (30 minutes). A written document plus a call to walk it through: where the money goes, which workloads drive it, what the realistic savings are, and what it would take to get them. It's yours whether or not you work with us after.
2. What we need from you
- Provider invoices or usage exports for the last 3 months, from whichever of OpenAI, Anthropic, AWS Bedrock, Google Vertex AI, or Azure OpenAI you use.
- Gateway or proxy logs if you have them (LiteLLM, Helicone, Langfuse, OpenRouter, or your own). Aggregated metadata only. If you don't have a gateway, that is itself part of the finding.
- A 10-minute description of your top workflows so token counts can be attributed to something a finance team recognizes.
Read-only access, no prompt or output content, no production credentials. The full scope is on the security page, and we'll sign an NDA before anything is exchanged.
3. What you get back
- Spend broken down by model, workload, and team, rather than one line per provider.
- The specific drivers: retries, oversized context, the wrong model on a cheap task, uncached repeats, agent loops nobody is watching.
- A ranked savings list with an estimated figure and an effort cost against each item.
- A baseline you can re-measure against later, so a future spike has something to be compared to.
4. Who this isn't for
We'd rather say this before the call than during it.
- Under $20k a month in AI spend. That's the minimum for a managed engagement — below it, a percentage of savings doesn't cover the work for either of us. The audit is still free if you want a second opinion, but our research and the cache savings calculator will get you most of the way on your own.
- Looking for a dashboard to buy. This is an engagement, not a product signup. If self-serve monitoring is what you want, that's a different tool.
- Unable to share billing data. There is no version of this that works from a description of the problem alone.
Questions people ask first
Is the audit really free?
Yes. It's how we qualify work, and it's cheaper for both of us than a proposal written blind. Most audits end with a specific list you could implement yourself. Some end with us doing it.
How long until we see the savings?
The audit takes about a week. The top one or two items on the list are usually configuration changes that land the same week you decide to make them. Anything requiring an architecture change is flagged as such.
Do we have to give you production access?
No. Read-only billing and usage exports are enough for the audit, and they're revocable by you at any time without asking us.
Who is on the call?
The person doing the audit. There is no separate sales stage to get past.