Everything runs on the largest model
The fix: routing
Picking the right model for the kind of request, escalating only when the answer falls short. Decided by measurement, not by feel.
Two starting points, one outcome: you know where you stand on AI today, what it costs and what to do next. Both assessments come at a fixed price, with no obligation.
You are just starting out
€4,900
Fixed price, net. A workshop with one pilot team, on your site. A standalone engagement, separate from the price ladder below.
For companies that want to adopt AI seriously and need to know where to start. We work with a real team on a real process, instead of laying a strategy across every department that nobody reads afterwards.
You already use AI
€3,000
Fixed price, net. A clear read on the token consumption of your AI systems.
For companies whose AI bill has been running for a while and who cannot break it down. It runs on your providers' billing data and three or four conversations.
A day or two with the people who actually run the process. Not with the people who would describe it.
What AI is already in the building, who uses it and for what. The shadow IT turns up almost every time.
EU AI Act risk class per use case, plus the gap against deployer obligations: oversight, logging, informing staff, AI literacy under Art. 4.
An order you can hold yourself to: which process first, what it needs, what it returns, and what would tell us it is not worth doing.
A day or two with the people who actually run the process. Not with the people who would describe it.
What AI is already in the building, who uses it and for what. The shadow IT turns up almost every time.
EU AI Act risk class per use case, plus the gap against deployer obligations: oversight, logging, informing staff, AI literacy under Art. 4.
An order you can hold yourself to: which process first, what it needs, what it returns, and what would tell us it is not worth doing.
What you spend on AI, how it has moved over the months and where it is heading.
Which model carries how much, and where the choice obviously does not match the task.
Reasoned, from billing data and conversations. Not a measurement, and we do not call it one.
The ones nobody in the building knew about. They turn up almost every time.
What you spend on AI, how it has moved over the months and where it is heading.
Which model carries how much, and where the choice obviously does not match the task.
Reasoned, from billing data and conversations. Not a measurement, and we do not call it one.
The ones nobody in the building knew about. They turn up almost every time.
A workflow running on the largest model when the smallest returns the same answer.
Repeated calls with identical input, because nothing is cached.
One team causing 60 % of the bill, without anyone having been able to name it.
The fix: routing
Picking the right model for the kind of request, escalating only when the answer falls short. Decided by measurement, not by feel.
The fix: caching
A layer in front of the model that recognises what has already been asked and reuses the answer.
The fix: rebuild the interaction
Often the model isn't too expensive, the dialogue is badly built. That is design work, not infrastructure.
The fix: an evaluation harness
Cost per result means nothing while nobody measures the quality of the result. We set up the harness that checks it.
Step 1 · Cost picture
€3,000
Fixed price, net. It runs on your providers' billing data and three or four conversations.
If the cost picture finds less than that, you get it refunded.
Applies from €10,000 of monthly AI spend
Step 2 · Measurement
10 % of your annual AI spend, €24,000 minimum
Inside your environment it is measured rather than estimated. The cost picture is credited in full if you commission within eight weeks.
| Annual spend | Measurement |
|---|---|
| €180,000 | €24,000 |
| €500,000 | €50,000 |
| €1,000,000 | €100,000 |
The cost picture and the guarantee do not scale with usage. The guarantee applies from €10,000 of monthly AI spend.
You know before every step what it costs. Analysis, prototype, implementation. All at a fixed price. New scope means a new mini-analysis and a new fixed price, not a quiet addendum. And because we can measure the benefit, we offer a success component if you want one: after twelve months, measured against your own baseline, with Tokenometrics as the shared data basis. The metric, the method and the attribution are agreed up front, not negotiated afterwards.
Tokenometrics is the platform you keep measuring with once we are gone.
Handover means handover: operations, documentation, runbooks and training sit with your team afterwards.
Tell us where you stand. Whether the bill has been running for a while or you are still before the first step. I'll get back to you within 24 hours.
