The price sheet lies

List price tells you almost nothing about what AI really costs.

Qwen3.6 27B (Non-reasoning) is listed 88% cheaper than GPT-5.5 (Non-reasoning). It costs 85% more to actually run.

The flip

Same models, two orderings

Each model ranked by list price on the left, and by what it actually costs to run Artificial Analysis's Intelligence Index on the right. Follow a line across — where it climbs, the sticker price lied.

LIST PRICE $ per million tokens ACTUAL COST $ to run the Intelligence Index HyperNova 60B 2605$0.07 HyperNova 60B 2605 — list $0.07/Mtok HY Qwen3.5 9B (Reasoning)$0.16 Qwen3.5 9B (Reasoning) — list $0.16/Mtok QW gpt-oss-120b (high)$0.26 gpt-oss-120b (high) — list $0.26/Mtok GP GPT-5.6 Luna (max)$0.45 GPT-5.6 Luna (max) — list $0.45/Mtok GT MiMo-V2.5-Pro$0.54 MiMo-V2.5-Pro — list $0.54/Mtok MI Mistral Large 3$0.75 Mistral Large 3 — list $0.75/Mtok MS Qwen3.5 122B A10B (Non-…$1.10 Qwen3.5 122B A10B (Non-reasoning) — list $1.10/Mtok QE Qwen3.6 27B (Non-reason…$1.35 Qwen3.6 27B (Non-reasoning) — list $1.35/Mtok QN Claude 4.5 Haiku (Reaso…$2.00 Claude 4.5 Haiku (Reasoning) — list $2.00/Mtok CL Gemini 3.5 Flash (high)$3.38 Gemini 3.5 Flash (high) — list $3.38/Mtok GE GPT-5 (high)$3.44 GPT-5 (high) — list $3.44/Mtok G5 Kimi K3 (low)$6.00 Kimi K3 (low) — list $6.00/Mtok KI GPT-5.5 (Non-reasoning)$11.25 GPT-5.5 (Non-reasoning) — list $11.25/Mtok G2 GPT-5.5 (xhigh)$11.25 GPT-5.5 (xhigh) — list $11.25/Mtok G3 $37HyperNova 60B 2605 HyperNova 60B 2605 — actually $37 to run the Intelligence Index ($0.0193/task) HY $241Qwen3.5 9B (Reasoning) Qwen3.5 9B (Reasoning) — actually $241 to run the Intelligence Index ($0.2211/task) QW $96gpt-oss-120b (high) gpt-oss-120b (high) — actually $96 to run the Intelligence Index ($0.0607/task) GP $191GPT-5.6 Luna (max) GPT-5.6 Luna (max) — actually $191 to run the Intelligence Index ($0.0571/task) GT $108MiMo-V2.5-Pro MiMo-V2.5-Pro — actually $108 to run the Intelligence Index ($0.0433/task) MI $71Mistral Large 3 Mistral Large 3 — actually $71 to run the Intelligence Index ($0.0599/task) MS $219Qwen3.5 122B A10B (No… Qwen3.5 122B A10B (Non-reasoning) — actually $219 to run the Intelligence Index ($0.1793/task) QE $357Qwen3.6 27B (Non-reas… Qwen3.6 27B (Non-reasoning) — actually $357 to run the Intelligence Index ($0.3602/task) QN $539Claude 4.5 Haiku (Rea… Claude 4.5 Haiku (Reasoning) — actually $539 to run the Intelligence Index ($0.2375/task) CL $1,041Gemini 3.5 Flash (hig… Gemini 3.5 Flash (high) — actually $1,041 to run the Intelligence Index ($0.5855/task) GE $793GPT-5 (high) GPT-5 (high) — actually $793 to run the Intelligence Index ($0.2215/task) G5 $283Kimi K3 (low) Kimi K3 (low) — actually $283 to run the Intelligence Index ($0.1885/task) KI $193GPT-5.5 (Non-reasonin… GPT-5.5 (Non-reasoning) — actually $193 to run the Intelligence Index ($0.1515/task) G2 $2,778GPT-5.5 (xhigh) GPT-5.5 (xhigh) — actually $2,778 to run the Intelligence Index ($0.9925/task) G3
Cheaper on paper, pricier to run Looks pricey, secretly a bargain Every other measured model

14 models with both a list price and a measured run cost · updated 8 hours ago

Why the sticker lies

You don't buy tokens. You buy answers.

A price sheet quotes dollars per million tokens. But a task isn't a fixed number of tokens — and that's where the bill hides.

List price is per token

The number on the pricing page is $ per million input/output tokens. It says nothing about how many tokens a model will spend to finish your task.

Reasoning models are verbose

A "thinking" model can emit 10–50× more tokens chewing through the same problem. Cheap per token × a mountain of tokens = an expensive answer.

The real unit is the task

Run the same fixed benchmark across models and the cheap-looking ones often cost the most. Rank by the finished job, not the sticker.

Receipts

Where list price misleads the most

Sorted by how far a model moves when you re-rank by real cost. ▲ means it's pricier than its sticker suggests; ▼ means it's a quiet bargain.

Model List $/Mtok Actual run cost Re-rank
Qwen3.5 9B (Reasoning) Alibaba $0.16 $241 ▲ 6 pricier
GPT-5.5 (Non-reasoning) OpenAI $11.25 $193 ▼ 7 cheaper
Mistral Large 3 Mistral $0.75 $71 ▼ 4 cheaper
Gemini 3.5 Flash (high) Google $3.38 $1,041 ▲ 3 pricier
Kimi K3 (low) Kimi $6.00 $283 ▼ 3 cheaper
Claude 4.5 Haiku (Reasoning) Anthropic $2.00 $539 ▲ 2 pricier
Qwen3.6 27B (Non-reasoning) Alibaba $1.35 $357 ▲ 2 pricier
GPT-5.6 Luna (max) OpenAI $0.45 $191 ▲ 1 pricier
GPT-5 (high) OpenAI $3.44 $793 ▲ 1 pricier
MiMo-V2.5-Pro Xiaomi $0.54 $108 ▼ 1 cheaper

List price is the blended 3:1 input:output rate. "Actual run cost" is what it costs to run the full Artificial Analysis Intelligence Index, from Artificial Analysis — a fixed task suite, so the only variable is how each model behaves.

Citation

Use this in your work

Every figure here is open data. Here's a ready-made citation, plus the methodology behind the numbers and the full State of AI.

Copy a citation

Free to use and cite under CC BY 4.0. See how this is measured.

APA

Champlin Enterprises. (2026). The AI Price Sheet Lies (MyTokenTracker) [Data set]. MyTokenTracker. Retrieved July 31, 2026, from https://mytokentracker.io/price-vs-cost

BibTeX
@misc{mytokentracker-price-vs-cost,
  title        = {The AI Price Sheet Lies (MyTokenTracker)},
  author       = {{Champlin Enterprises}},
  year         = {2026},
  howpublished = {MyTokenTracker, \url{https://mytokentracker.io/price-vs-cost}},
  note         = {Accessed July 31, 2026. Licensed CC BY 4.0.},
  url          = {https://mytokentracker.io/price-vs-cost}
}

Need a fixed point in time? Every day’s data is permanently archived in the open-data repository, so you can cite a specific date by linking that day’s committed file.

Free weekly digest

Stop guessing what AI costs

We track the real, measured cost of every model — not the sticker price. One line to install, free forever. Get the weekly headline in your inbox.

No spam, no account. One click to leave.