Calculator
What will this actually cost?
Most pricing pages multiply tokens in by a rate and tokens out by another. On anything agentic that is the small half of the bill. Set a workload below and this prices the prompt cache too — reads and writes.
Why the cache write is the line that matters
A prompt cache lets a vendor skip re-processing a prefix you have sent before, and charges you much less for those tokens — often a tenth of the input rate. The part that is easy to miss is that putting something into the cache costs more than sending it normally: vendors that publish the figure charge around 1.25× the input rate for a five-minute entry and 2× for an hour.
So a workload that rewrites its cache often can cost more than one that never caches at all. That is not an edge case — it is what happens when an agent edits its own context between turns. Set the cache writes above to a high number and watch the total move.