Not available for this agent
Priced by Sentry
$51.72
List-equivalent, not billed
Floor on this volume
$1.91
All of it on gpt-5.4-nano
Headroom
$49.81
The gap routing could close, at most
Cache hit rate
100%
72.8M tokens read from cache
What this means
This agent moved 77.3M tokens across 1 model, priced at $51.72. The same token volume on gpt-5.4-nano would price at $1.91. The difference, $49.81, is the most routing could ever recover, not a saving on offer. Whether the cheaper model would have done the job is a judgement telemetry cannot make for you, and any vendor quoting you the full number as a saving is selling you arithmetic.
Note: this agent is on a seat, so none of this is money anyone was billed. The figures are list-equivalent. They still matter, because they tell you what the same work would cost the day you move it onto the API.
Where these levers actually come from
Where the tokens went
What to do about it
One model is doing everything
Every call on this agent went to claude-sonnet-5. That is simple, and it is also the most expensive way to run an agent. Splitting the easy calls onto a smaller model is where the first real reduction comes from.