What AI actually costs, right now.
One live, citable view of the whole picture: the AI Cost Index, what our community really spends, the biggest price moves, and every model price you can filter and sort yourself. Built from real usage. Free under CC BY 4.0.
0
models priced
0
providers tracked
$0
community spend tracked
0
tokens tracked
Prices updated 9 hours ago · refreshes automatically
Key findings
What the numbers say right now
GPT-5.6 Luna (max) reaches 85% of the top intelligence score at 5% of the price
Versus Claude Opus 5 (Adaptive Reasoning, Xhigh Effort), the current intelligence leader. Quality from Artificial Analysis; price blended 3:1.
One dollar now buys about 61,538 pages of AI-written text
At Gemma 3n E4B Instruct, one of the cheapest capable models, $0.03 per million tokens (blended).
A 100,000-word novel costs about $3.25 to generate on Claude Opus 5 (Adaptive Reasoning, Xhigh Effort)
Roughly 130,000 output tokens on today's top-rated model.
The priciest model costs about 10,500x more than the cheapest for the same million tokens
o1-pro vs Gemma 3n E4B Instruct, blended price. Picking the right model is the biggest cost lever you have.
Open-weight models now run as low as $0.06 per million tokens
Qwen3.5 4B (Non-reasoning) is the cheapest open-weight model tracked, roughly 25,641 pages per dollar.
Computed live from our price, quality, and index data. Free to use and cite under CC BY 4.0.
Benchmark
The AI Cost Index
What a million tokens costs across a fixed basket of models, tracked over time. The S&P 500 of AI prices.
Frontier
One flagship model per major provider
$4.64/Mtok
0% since start
Budget
One cost-efficient workhorse per major provider
$0.7/Mtok
0% since start
Price war
Biggest recent price moves
When a provider changes a price, it shows up here. Green is a cut, red is a hike (blended $/Mtok).
mistral/mistral-medium-latest
Mistral · $0.8 → $3
sambanova/MiniMax-M2.7
Sambanova · $0.52 → $1.05
sambanova/gpt-oss-120b
Sambanova · $3.38 → $0.31
amazon.titan-embed-text-v2:0
Bedrock · $0.15 → $0.02
bedrock_mantle/openai.gpt-5.6-luna
Bedrock Mantle · $2.48 → $0.5
gpt-5.6-luna
Openai · $2.25 → $0.45
Intelligence & speed
Not just cheap, capable
Independent quality and speed benchmarks, so cost reads next to capability. Smartest, fastest, and the best capability per dollar.
🧠 Smartest
Intelligence Index
- 1 Claude Opus 5 (Adaptive Reasoning, Xhigh Effort) Anthropic 60.1
- 2 GPT-5.5 (xhigh) OpenAI 54.8
- 3 Claude Sonnet 5 (Adaptive Reasoning, Max Effort) Anthropic 53.4
- 4 GPT-5.6 Luna (max) OpenAI 51.2
- 5 Claude Opus 5 (Adaptive Reasoning, Low Effort) Anthropic 50.6
- 6 GPT-5.5 (medium) OpenAI 50.4
⚡ Fastest
Output speed
- 1 Mercury 2 Inception 945 tok/s
- 2 Step 3.7 Flash StepFun 413 tok/s
- 3 LFM2.5-VL-1.6B Liquid AI 401 tok/s
- 4 HyperNova 60B 2605 Multiverse Computing 375 tok/s
- 5 Gemini 3.5 Flash-Lite Google 368 tok/s
- 6 Granite 3.3 8B (Non-reasoning) IBM 368 tok/s
💎 Best value
Intelligence per $/Mtok
- 1 HyperNova 60B 2605 Multiverse Computing 273.8 pts/$
- 2 Qwen3.5 4B (Non-reasoning) Alibaba 266.7 pts/$
- 3 Gemma 4 E4B (Non-reasoning) Google 222.5 pts/$
- 4 DeepSeek V4 Flash (Reasoning, High Effort) DeepSeek 213.7 pts/$
- 5 MiMo-V2.5 Xiaomi 212.6 pts/$
- 6 Step 3.5 Flash 2603 StepFun 173.3 pts/$
Quality & speed data from Artificial Analysis.
Human preference
What people actually prefer
The LMArena (Chatbot Arena) overall leaderboard, decided by millions of blind head-to-head votes. Snapshot Jul 21, 2026.
- 1 claude-opus-4-6-thinking Anthropic 1503 63.2K votes
- 2 gpt-5.4 Openai 1499 61.8K votes
- 3 claude-opus-4-6 Anthropic 1497 67K votes
- 4 gpt-5.4-mini-high Openai 1496 57.6K votes
- 5 claude-fable-5 Anthropic 1493 14.6K votes
- 6 gpt-5.2-high Openai 1493 47.4K votes
- 7 gemini-3.6-flash Google 1491 4.7K votes
- 8 gemini-3.5-flash-high Google 1490 10.1K votes
- 9 claude-opus-4-7-thinking Anthropic 1490 50.7K votes
- 10 gpt-5.5 Openai 1488 47K votes
Data: LMArena leaderboard, CC BY 4.0.
Real usage
What the community really spends
Anonymized, aggregate, opt-in. Numbers from real engineers, not list prices.
Tracked spend
$648.15
Tokens
67B
Tokens / $
103.3M
Contributors
5
Daily community spend
By provider
-
Anthropic 100%
By platform
-
Claude Code 100%
By use case
-
Coding 100%
Value
Cheapest & priciest right now
Blended $/Mtok (3:1 input:output) across text models with public prices.
Cheapest
- fireworks_ai/accounts/fireworks/models/flux-1-dev-controlnet-union Fireworks Ai $0/Mtok
- fireworks-ai-embedding-up-to-150m Fireworks Ai Embedding Models $0.01/Mtok
- fireworks-ai-embedding-150m-to-350m Fireworks Ai Embedding Models $0.01/Mtok
- nscale/Qwen/Qwen2.5-Coder-3B-Instruct Nscale $0.02/Mtok
- nscale/Qwen/Qwen2.5-Coder-7B-Instruct Nscale $0.02/Mtok
- nebius/Qwen/Qwen2.5-Coder-7B Nebius $0.02/Mtok
Priciest
- wandb/deepseek-ai/DeepSeek-R1-0528 Wandb $236,250/Mtok
- wandb/deepseek-ai/DeepSeek-V3-0324 Wandb $154,250/Mtok
- wandb/Qwen/Qwen3-Coder-480B-A35B-Instruct Wandb $112,500/Mtok
- wandb/zai-org/GLM-4.5 Wandb $91,250/Mtok
- wandb/deepseek-ai/DeepSeek-V3.1 Wandb $82,500/Mtok
- wandb/meta-llama/Llama-3.3-70B-Instruct Wandb $71,000/Mtok
Explore
Every model, your way
Filter by provider, search, and sort the entire live price catalog. Pulled straight from the open Data API.
| Model | Provider | Input | Output | Blended |
|---|
Source: GET /api/v1/data/models · CC BY 4.0
The cheatsheet
Which AI for which job
No single model wins everything. The top pick for each job, ranked from real data.
Reasoning
Claude Opus 5 (Adaptive Reasoning, Xhigh Effort)
Coding
Claude Opus 5 (Adaptive Reasoning, Xhigh Effort)
Agents
Claude Opus 5 (Adaptive Reasoning, Xhigh Effort)
Chat & writing
claude-opus-4-6-thinking
Speed
Step 3.7 Flash
Long context
Llama 4 Scout 17b 128e Instruct Maas
Best value
Qwen3.5 4B (Non-reasoning)
Budget
Llama 3.1 8b
The price sheet lies
List price isn't what you pay
Qwen3.6 27B (Non-reasoning) is listed 88% cheaper than GPT-5.5 (Non-reasoning) — yet costs 85% more to actually run. Each model by list price (left) vs the real bill to run the Intelligence Index (right).
Run cost = cost to run the full Intelligence Index, from Artificial Analysis.
Citation
Use this in your work
Quote any figure on this page. It is all open data. Here is a ready-made citation, and the methodology behind every number.
Copy a citation
Free to use and cite under CC BY 4.0. See how this is measured.
Champlin Enterprises. (2026). The State of AI (MyTokenTracker) [Data set]. MyTokenTracker. Retrieved July 31, 2026, from https://mytokentracker.io/state-of-ai
@misc{mytokentracker-state-of-ai,
title = {The State of AI (MyTokenTracker)},
author = {{Champlin Enterprises}},
year = {2026},
howpublished = {MyTokenTracker, \url{https://mytokentracker.io/state-of-ai}},
note = {Accessed July 31, 2026. Licensed CC BY 4.0.},
url = {https://mytokentracker.io/state-of-ai}
}
Need a fixed point in time? Every day’s data is permanently archived in the open-data repository, so you can cite a specific date by linking that day’s committed file.
Free weekly digest
Get this in your inbox, weekly
The State of AI moves every day. Once a week we send the headline: the AI Cost Index, the biggest price moves, and what it means. Free, no account.
Measure it. Don't guess.
Every number on this page came from people who track instead of guess. One line to install, automatic capture, free forever. Add your usage and the whole picture gets sharper.
Our wiggly friend is an original MyTokenTracker mascot, here purely for fun. It is not affiliated with, endorsed by, or representing Anthropic, OpenAI, Google, or any model provider. Product and model names (Claude, GPT, Gemini, and others) are trademarks of their respective owners and appear here only to report public pricing and usage.