What AI actually costs, right now.
One live, citable view of the whole picture: the AI Cost Index, what our community really spends, the biggest price moves, and every model price you can filter and sort yourself. Built from real usage. Free under CC BY 4.0.
0
models priced
0
providers tracked
$0
community spend tracked
0
tokens tracked
Prices updated 15 hours ago · refreshes automatically
Key findings
What the numbers say right now
Gemini 3.5 Flash (medium) reaches 87% of the top intelligence score at 17% of the price
Versus Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback), the current intelligence leader. Quality from Artificial Analysis; price blended 3:1.
One dollar now buys about 61,538 pages of AI-written text
At Gemma 3n E4B Instruct, one of the cheapest capable models, $0.03 per million tokens (blended).
A 100,000-word novel costs about $6.50 to generate on Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback)
Roughly 130,000 output tokens on today's top-rated model.
The priciest model costs about 10,500x more than the cheapest for the same million tokens
o1-pro vs Gemma 3n E4B Instruct, blended price. Picking the right model is the biggest cost lever you have.
Open-weight models now run as low as $0.06 per million tokens
Qwen3.5 4B (Non-reasoning) is the cheapest open-weight model tracked, roughly 25,641 pages per dollar.
Computed live from our price, quality, and index data. Free to use and cite under CC BY 4.0.
Benchmark
The AI Cost Index
What a million tokens costs across a fixed basket of models, tracked over time. The S&P 500 of AI prices.
Frontier
One flagship model per major provider
$4.64/Mtok
0% since start
Budget
One cost-efficient workhorse per major provider
$0.74/Mtok
+4.9% since start
Price war
Biggest recent price moves
When a provider changes a price, it shows up here. Green is a cut, red is a hike (blended $/Mtok).
xai/grok-4-fast-reasoning
Xai · $0.28 → $1.56
xai/grok-4-fast-non-reasoning
Xai · $0.28 → $1.56
xai/grok-4-1-fast
Xai · $0.28 → $1.56
xai/grok-4-1-fast-reasoning
Xai · $0.28 → $1.56
xai/grok-4-1-fast-reasoning-latest
Xai · $0.28 → $1.56
xai/grok-4-1-fast-non-reasoning
Xai · $0.28 → $1.56
Intelligence & speed
Not just cheap, capable
Independent quality and speed benchmarks, so cost reads next to capability. Smartest, fastest, and the best capability per dollar.
🧠 Smartest
Intelligence Index
- 1 Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback) Anthropic 53.4
- 2 GPT-6 Astra (max) OpenAI 52.8
- 3 GPT-6 Astra (xhigh) OpenAI 52.5
- 4 GPT-5.5 (medium) OpenAI 50.4
- 5 Claude Opus 5 (Adaptive Reasoning, Xhigh Effort) Anthropic 49.7
- 6 Claude Fable 5.1 (Adaptive Reasoning, Medium Effort, Default Fallback) Anthropic 49.1
⚡ Fastest
Output speed
- 1 Mercury 2 Inception 807 tok/s
- 2 Step 3.7 Flash StepFun 413 tok/s
- 3 HyperNova 60B 2605 (high, based on gpt-oss-120b) Multiverse Computing 361 tok/s
- 4 gpt-oss-120b (low) OpenAI 346 tok/s
- 5 gpt-oss-120b (high) OpenAI 336 tok/s
- 6 Gemini 3.1 Flash-Lite Google 333 tok/s
💎 Best value
Intelligence per $/Mtok
- 1 Gemma 4 E4B (Non-reasoning) Google 217.5 pts/$
- 2 DeepSeek V4 Flash (Reasoning, High Effort) DeepSeek 213.7 pts/$
- 3 Qwen3.5 4B (Non-reasoning) Alibaba 180 pts/$
- 4 HyperNova 60B 2605 (high, based on gpt-oss-120b) Multiverse Computing 180 pts/$
- 5 GLM-5.3-Flash Z AI 176.4 pts/$
- 6 Granite 4.2 3B IBM 173.3 pts/$
Quality & speed data from Artificial Analysis.
Human preference
What people actually prefer
The LMArena (Chatbot Arena) overall leaderboard, decided by millions of blind head-to-head votes. Snapshot Sep 13, 2026.
- 1 claude-fable-5.1-max Anthropic 1508 5.8K votes
- 2 claude-opus-5-max Anthropic 1505 20.7K votes
- 3 claude-opus-5-high Anthropic 1505 42.6K votes
- 4 claude-opus-4-6-high Anthropic 1503 72K votes
- 5 claude-opus-4-6 Anthropic 1498 75.9K votes
- 6 gemini-3.8-flash-high Google 1495 5.1K votes
- 7 claude-fable-5 Anthropic 1493 30.1K votes
- 8 gemini-3.7-flash-high Google 1491 5.6K votes
- 9 claude-opus-4-7-high Anthropic 1490 60K votes
- 10 muse-spark-1.3-max Meta 1490 4.7K votes
Data: LMArena leaderboard, CC BY 4.0.
Real usage
What the community really spends
Anonymized, aggregate, opt-in. Numbers from real engineers, not list prices.
Tracked spend
$1,104.03
Tokens
128.2B
Tokens / $
116.1M
Contributors
6
Daily community spend
By provider
-
Anthropic 100%
By platform
-
Claude Code 100%
By use case
-
Coding 100%
Value
Cheapest & priciest right now
Blended $/Mtok (3:1 input:output) across text models with public prices.
Cheapest
- fireworks_ai/accounts/fireworks/models/flux-1-dev-controlnet-union Fireworks Ai $0/Mtok
- fireworks-ai-embedding-up-to-150m Fireworks Ai Embedding Models $0.01/Mtok
- fireworks-ai-embedding-150m-to-350m Fireworks Ai Embedding Models $0.01/Mtok
- nscale/Qwen/Qwen2.5-Coder-3B-Instruct Nscale $0.02/Mtok
- nscale/Qwen/Qwen2.5-Coder-7B-Instruct Nscale $0.02/Mtok
- nebius/Qwen/Qwen2.5-Coder-7B Nebius $0.02/Mtok
Priciest
- wandb/zai-org/GLM-4.5 Wandb $91,250/Mtok
- wandb/microsoft/Phi-4-mini-instruct Wandb $14,750/Mtok
- azure_ai/jais-30b-chat Azure $4,827.5/Mtok
- watsonx/core42/jais-13b-chat Watsonx $875/Mtok
- openrouter/openai/o1-pro Openrouter $262.5/Mtok
- o1-pro-2025-03-19 Openai $262.5/Mtok
Explore
Every model, your way
Filter by provider, search, and sort the entire live price catalog. Pulled straight from the open Data API.
| Model | Provider | Input | Output | Blended |
|---|
Source: GET /api/v1/data/models · CC BY 4.0
The cheatsheet
Which AI for which job
No single model wins everything. The top pick for each job, ranked from real data.
Reasoning
Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback)
Coding
Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback)
Agents
Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback)
Chat & writing
claude-fable-5.1-max
Speed
Step 3.7 Flash
Long context
Llama 4 Scout 17b 128e Instruct Maas
Best value
Gemma 4 E4B (Non-reasoning)
Budget
Llama 3.1 8b
The price sheet lies
List price isn't what you pay
MiniMax-M3 is listed 91% cheaper than Kimi K3 (low) — yet costs 90% more to actually run. Each model by list price (left) vs the real bill to run the Intelligence Index (right).
Run cost = cost to run the full Intelligence Index, from Artificial Analysis.
Citation
Use this in your work
Quote any figure on this page. It is all open data. Here is a ready-made citation, and the methodology behind every number.
Copy a citation
Free to use and cite under CC BY 4.0. See how this is measured.
Champlin Enterprises. (2026). The State of AI (MyTokenTracker) [Data set]. MyTokenTracker. Retrieved September 15, 2026, from https://mytokentracker.io/state-of-ai
@misc{mytokentracker-state-of-ai,
title = {The State of AI (MyTokenTracker)},
author = {{Champlin Enterprises}},
year = {2026},
howpublished = {MyTokenTracker, \url{https://mytokentracker.io/state-of-ai}},
note = {Accessed September 15, 2026. Licensed CC BY 4.0.},
url = {https://mytokentracker.io/state-of-ai}
}
Need a fixed point in time? Every day’s data is permanently archived in the open-data repository, so you can cite a specific date by linking that day’s committed file.
Free weekly digest
Get this in your inbox, weekly
The State of AI moves every day. Once a week we send the headline: the AI Cost Index, the biggest price moves, and what it means. Free, no account.
Measure it. Don't guess.
Every number on this page came from people who track instead of guess. One line to install, automatic capture, free forever. Add your usage and the whole picture gets sharper.
Our wiggly friend is an original MyTokenTracker mascot, here purely for fun. It is not affiliated with, endorsed by, or representing Anthropic, OpenAI, Google, or any model provider. Product and model names (Claude, GPT, Gemini, and others) are trademarks of their respective owners and appear here only to report public pricing and usage.