Vibedia.
All vendors

Show only the vendors you work with. This applies to every page, and is remembered.

Learn · 4 of 7

What it costs, and why the surprise

Where the money actually goes, which is rarely where people expect.

Prices are quoted per million tokens (How model prices are quoted: dollars per million tokens, charged separately for what you send and what comes back.), in each direction, and a million tokens is less than it sounds: roughly one mid-sized codebase, or one long afternoon with a coding agent.

The spread between models is about six hundred times from cheapest to dearest, which is the first surprise. The second is where your own money actually goes.

Three things dominate, in this order

Context, re-sent. Nothing persists between calls, so every turn re-sends the whole conversation. A 100,000-token codebase across forty turns is four million input tokens. The answers in the same session might total eighty thousand. The context is fifty times the output.

Cache writes (The charge for putting a prefix into the prompt cache. Usually more than the input rate, not less.). If you are caching that context — and you should be — writing an entry costs more than sending it normally: around twice the input rate for a one-hour entry. Rewrite it every turn and you pay the premium every turn.

Output, last. It has the highest rate per token, but on most real workloads there is far less of it than people assume.

That ordering is why “use a cheaper model” is often the wrong first move, and “send less context” is usually the right one.

The model that costs nothing extra

Training a frontier model (The largest and most capable model a vendor currently offers, priced accordingly.) costs tens or hundreds of millions of dollars. You are never charged for any of it. What you pay for is inference (Running a trained model to get an answer. Everything you pay for by the token is inference, not training.) — running the finished model — and no amount of your usage teaches it anything, because the weights are fixed between releases.

Working it out rather than guessing

The calculator takes a real workload — tokens, calls a day, cache behaviour — and prices it against every model at once. Where a vendor does not publish a figure it needs, it refuses to produce a number rather than filling the gap, which is the whole point.

The spread

Input price per million tokens, on a logarithmic scale

Input price per million tokens, by vendor64 models from 7 vendors plotted on a logarithmic price axis. The cheapest is GPT-5 Nano at $0.05 and the dearest is GPT-5.5 Pro at $30, a spread of about 600 times. The same figures are in the table below.$0.1$0.3$1$3$10$30OpenAI GPT-3.5-turbo: $0.5 in, $1.5 out GPT-4: $30 in, $60 out GPT-4 Turbo: $10 in, $30 out GPT-4.1: $2 in, $8 out GPT-4.1 mini: $0.4 in, $1.6 out GPT-4.1 nano: $0.1 in, $0.4 out GPT-4o: $2.5 in, $10 out GPT-4o mini: $0.15 in, $0.6 out GPT-5: $1.25 in, $10 out GPT-5 Mini: $0.25 in, $2 out GPT-5 Nano: $0.05 in, $0.4 out GPT-5 Pro: $15 in, $120 out GPT-5.1: $1.25 in, $10 out GPT-5.2: $1.75 in, $14 out GPT-5.2 Pro: $21 in, $168 out GPT-5.3 Codex: $1.75 in, $14 out GPT-5.4: $2.5 in, $15 out GPT-5.4 Mini: $0.75 in, $4.5 out GPT-5.4 Nano: $0.2 in, $1.25 out GPT-5.4 Pro: $30 in, $180 out GPT-5.5: $5 in, $30 out GPT-5.5 Pro: $30 in, $180 out GPT-5.6: $4 in, $20 out GPT-5.6 Luna: $0.2 in, $1.2 out GPT-5.6 Sol: $4 in, $20 out GPT-5.6 Terra: $2 in, $12 out gpt-5.6-cyber: $12.5 in, $75 out GPT-6 Astra: $10 in, $50 out GPT-6 Luna: $0.1 in, $0.5 out GPT-6 Sol: $2 in, $10 out GPT-6.1 Sol: $2 in, $10 out gpt-rosalind-discovery: $5 in, $25 out gpt-rosalind-research: $5 in, $25 out Anthropic Claude Fable 5: $10 in, $50 out Claude Fable 5.1: $10 in, $50 out Claude Haiku 4.5: $1 in, $5 out Claude Haiku 5.5: $0.1 in, $0.5 out Claude Mythos 5.1: $10 in, $50 out Claude Opus 4.5: $5 in, $25 out Claude Opus 4.6: $5 in, $25 out Claude Opus 4.7: $5 in, $25 out Claude Opus 4.8: $5 in, $25 out Claude Opus 5: $5 in, $25 out Claude Opus 5.5: $4 in, $20 out Claude Sonnet 4.5 (latest): $3 in, $15 out Claude Sonnet 4.6: $3 in, $15 out Claude Sonnet 5: $2 in, $10 out Claude Sonnet 5.5: $2 in, $10 out Google Gemini 2.5 Flash: $0.3 in, $2.5 out Gemini 2.5 Flash Lite: $0.1 in, $0.4 out Gemini 2.5 Pro: $1.25 in, $10 out Gemini 3 Flash Preview: $0.5 in, $3 out Gemini 3.1 Flash Lite: $0.25 in, $1.5 out Gemini 3.1 Flash Live Preview: $0.75 in, $4.5 out Gemini 3.1 Pro Preview: $2 in, $12 out Gemini 3.5 Flash Lite: $0.3 in, $2.5 out Gemini 3.6 Flash: $0.75 in, $3.75 out Gemini 3.8 Flash: $0.75 in, $3.75 out Gemini Omni Flash Preview: $1.5 in, $17.5 out AI21 Jamba Large: $2 in, $8 out Jamba Mini: $0.2 in, $0.4 out Cohere Command R+: $2.5 in, $10 out Mistral AI Codestral 25.08: $0.3 in, $0.9 out xAI Grok 4.20 (Non-Reasoning): $1.25 in, $2.5 out
64 models. The cheapest input price isGPT-5 Nano at $0.05; the dearest is GPT-5.5 Pro at $30 — about 600 times more for the same million tokens. Output prices, which are usually several times higher, are in the table.
The same figures as a table
ModelVendorInputOutput
GPT-5 NanoOpenAI$0.05$0.4
Claude Haiku 5.5Anthropic$0.1$0.5
GPT-6 LunaOpenAI$0.1$0.5
Gemini 2.5 Flash LiteGoogle$0.1$0.4
GPT-4.1 nanoOpenAI$0.1$0.4
GPT-4o miniOpenAI$0.15$0.6
GPT-5.6 LunaOpenAI$0.2$1.2
GPT-5.4 NanoOpenAI$0.2$1.25
Jamba MiniAI21$0.2$0.4
GPT-5 MiniOpenAI$0.25$2
Gemini 3.1 Flash LiteGoogle$0.25$1.5
Gemini 3.5 Flash LiteGoogle$0.3$2.5
Gemini 2.5 FlashGoogle$0.3$2.5
Codestral 25.08Mistral AI$0.3$0.9
GPT-4.1 miniOpenAI$0.4$1.6
Gemini 3 Flash PreviewGoogle$0.5$3
GPT-3.5-turboOpenAI$0.5$1.5
GPT-5.4 MiniOpenAI$0.75$4.5
Gemini 3.8 FlashGoogle$0.75$3.75
Gemini 3.6 FlashGoogle$0.75$3.75
Gemini 3.1 Flash Live PreviewGoogle$0.75$4.5
Claude Haiku 4.5Anthropic$1$5
GPT-5.1OpenAI$1.25$10
GPT-5OpenAI$1.25$10
Gemini 2.5 ProGoogle$1.25$10
Grok 4.20 (Non-Reasoning)xAI$1.25$2.5
Gemini Omni Flash PreviewGoogle$1.5$17.5
GPT-5.2OpenAI$1.75$14
GPT-5.3 CodexOpenAI$1.75$14
Claude Sonnet 5.5Anthropic$2$10
Claude Sonnet 5Anthropic$2$10
GPT-6.1 SolOpenAI$2$10
GPT-6 SolOpenAI$2$10
GPT-5.6 TerraOpenAI$2$12
Gemini 3.1 Pro PreviewGoogle$2$12
Jamba LargeAI21$2$8
GPT-4.1OpenAI$2$8
GPT-5.4OpenAI$2.5$15
Command R+Cohere$2.5$10
GPT-4oOpenAI$2.5$10
Claude Sonnet 4.6Anthropic$3$15
Claude Sonnet 4.5 (latest)Anthropic$3$15
Claude Opus 5.5Anthropic$4$20
GPT-5.6 SolOpenAI$4$20
GPT-5.6OpenAI$4$20
Claude Opus 5Anthropic$5$25
Claude Opus 4.8Anthropic$5$25
Claude Opus 4.7Anthropic$5$25
Claude Opus 4.6Anthropic$5$25
Claude Opus 4.5Anthropic$5$25
GPT-5.5OpenAI$5$30
gpt-rosalind-discoveryOpenAI$5$25
gpt-rosalind-researchOpenAI$5$25
Claude Fable 5.1Anthropic$10$50
Claude Fable 5Anthropic$10$50
Claude Mythos 5.1Anthropic$10$50
GPT-6 AstraOpenAI$10$50
GPT-4 TurboOpenAI$10$30
gpt-5.6-cyberOpenAI$12.5$75
GPT-5 ProOpenAI$15$120
GPT-5.2 ProOpenAI$21$168
GPT-5.5 ProOpenAI$30$180
GPT-4OpenAI$30$60
GPT-5.4 ProOpenAI$30$180
ShareOpen LinkedIn