Live · The State of AI

What AI actually costs, right now.

One live, citable view of the whole picture: the AI Cost Index, what our community really spends, the biggest price moves, and every model price you can filter and sort yourself. Built from real usage. Free under CC BY 4.0.

0

models priced

0

providers tracked

$0

community spend tracked

0

tokens tracked

Prices updated 20 hours ago · refreshes automatically

fireworks_ai/accounts/fireworks/models/flux-1-dev-controlnet-union $0 /Mtok qwencloud/qwen-turbo-2025-04-28 $0.09 /Mtok ovhcloud/Qwen3-32B $0.12 /Mtok nebius/google/gemma-3-27b-it $0.15 /Mtok perplexity/llama-3.1-8b-instruct $0.2 /Mtok wandb/meta-llama/Llama-3.1-8B-Instruct $0.22 /Mtok azure/gpt-4o-mini $0.29 /Mtok novita/qwen/qwen-2.5-72b-instruct $0.39 /Mtok anthropic.claude-3-haiku-20240307-v1:0 $0.5 /Mtok dashscope/qwen-plus-2025-01-25 $0.6 /Mtok qwen_ai_platform/qwen3-vl-235b-a22b-instruct $0.7 /Mtok minimax/MiniMax-M2.1-lightning $0.83 /Mtok fireworks_ai/accounts/fireworks/models/qwen2p5-coder-32b-instruct-64k $0.9 /Mtok fireworks_ai/accounts/fireworks/models/kimi-k2-instruct $1.08 /Mtok tensormesh/Qwen/Qwen3.5-397B-A17B-FP8 $1.35 /Mtok xai/grok-4.20-reasoning $1.56 /Mtok novita/deepseek/deepseek-v4-pro-0813 $1.98 /Mtok moonshot/kimi-latest $2.75 /Mtok azure/gpt-5-codex $3.44 /Mtok openrouter/anthropic/claude-sonnet-5 $4 /Mtok azure/gpt-5.4 $5.63 /Mtok openrouter/anthropic/claude-sonnet-4.5 $6 /Mtok azure_ai/claude-opus-5 $10 /Mtok vertex_ai/claude-fable-5-1@default $20 /Mtok fireworks_ai/accounts/fireworks/models/flux-1-dev-controlnet-union $0 /Mtok qwencloud/qwen-turbo-2025-04-28 $0.09 /Mtok ovhcloud/Qwen3-32B $0.12 /Mtok nebius/google/gemma-3-27b-it $0.15 /Mtok perplexity/llama-3.1-8b-instruct $0.2 /Mtok wandb/meta-llama/Llama-3.1-8B-Instruct $0.22 /Mtok azure/gpt-4o-mini $0.29 /Mtok novita/qwen/qwen-2.5-72b-instruct $0.39 /Mtok anthropic.claude-3-haiku-20240307-v1:0 $0.5 /Mtok dashscope/qwen-plus-2025-01-25 $0.6 /Mtok qwen_ai_platform/qwen3-vl-235b-a22b-instruct $0.7 /Mtok minimax/MiniMax-M2.1-lightning $0.83 /Mtok fireworks_ai/accounts/fireworks/models/qwen2p5-coder-32b-instruct-64k $0.9 /Mtok fireworks_ai/accounts/fireworks/models/kimi-k2-instruct $1.08 /Mtok tensormesh/Qwen/Qwen3.5-397B-A17B-FP8 $1.35 /Mtok xai/grok-4.20-reasoning $1.56 /Mtok novita/deepseek/deepseek-v4-pro-0813 $1.98 /Mtok moonshot/kimi-latest $2.75 /Mtok azure/gpt-5-codex $3.44 /Mtok openrouter/anthropic/claude-sonnet-5 $4 /Mtok azure/gpt-5.4 $5.63 /Mtok openrouter/anthropic/claude-sonnet-4.5 $6 /Mtok azure_ai/claude-opus-5 $10 /Mtok vertex_ai/claude-fable-5-1@default $20 /Mtok

Key findings

What the numbers say right now

85% quality, 15% price

Grok 4.6 (medium) reaches 85% of the top intelligence score at 15% of the price

Versus Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback), the current intelligence leader. Quality from Artificial Analysis; price blended 3:1.

61,538 pages for $1

One dollar now buys about 61,538 pages of AI-written text

At Gemma 3n E4B Instruct, one of the cheapest capable models, $0.03 per million tokens (blended).

$6.50

A 100,000-word novel costs about $6.50 to generate on Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback)

Roughly 130,000 output tokens on today's top-rated model.

10,500x spread

The priciest model costs about 10,500x more than the cheapest for the same million tokens

o1-pro vs Gemma 3n E4B Instruct, blended price. Picking the right model is the biggest cost lever you have.

$0.06 / 1M

Open-weight models now run as low as $0.06 per million tokens

Qwen3.5 4B (Non-reasoning) is the cheapest open-weight model tracked, roughly 25,641 pages per dollar.

Computed live from our price, quality, and index data. Free to use and cite under CC BY 4.0.

Benchmark

The AI Cost Index

What a million tokens costs across a fixed basket of models, tracked over time. The S&P 500 of AI prices.

Frontier

One flagship model per major provider

$4.64/Mtok

0% since start

$3.40 $4.23 $5.05 $5.88 Jun 15: $4.64 Jun 16: $4.64 Jun 17: $4.64 Jun 19: $4.64 Jun 20: $4.64 Jun 21: $4.64 Jun 22: $4.64 Jun 23: $4.64 Jun 24: $4.64 Jun 25: $4.64 Jun 26: $4.64 Jun 27: $4.64 Jun 28: $4.64 Jun 29: $4.64 Jun 30: $4.64 Jul 1: $4.64 Jul 2: $4.64 Jul 3: $4.64 Jul 4: $4.64 Jul 5: $4.64 Jul 6: $4.64 Jul 7: $4.64 Jul 8: $4.64 Jul 9: $4.64 Jul 10: $4.64 Jul 11: $4.64 Jul 12: $4.64 Jul 13: $4.64 Jul 14: $4.64 Jul 15: $4.64 Jul 16: $4.64 Jul 17: $4.64 Jul 18: $4.64 Jul 19: $4.64 Jul 20: $4.64 Jul 21: $4.64 Jul 22: $4.64 Jul 23: $4.64 Jul 24: $4.64 Jul 25: $4.64 Jul 26: $4.64 Jul 27: $4.64 Jul 28: $4.64 Jul 29: $4.64 Jul 30: $4.64 Jul 31: $4.64 Aug 1: $4.64 Aug 2: $4.64 Aug 3: $4.64 Aug 4: $4.64 Aug 6: $4.64 Aug 8: $4.64 Aug 9: $4.64 Aug 10: $4.64 Aug 11: $4.64 Aug 12: $4.64 Aug 13: $4.64 Aug 14: $4.64 Aug 15: $4.64 Aug 16: $4.64 Aug 17: $4.64 Aug 18: $4.64 Aug 19: $4.64 Aug 21: $4.64 Aug 22: $4.64 Aug 23: $4.64 Aug 24: $4.64 Aug 25: $4.64 Aug 26: $4.64 Aug 27: $4.64 Aug 28: $4.64 Aug 29: $4.64 Aug 30: $4.64 Aug 31: $4.64 Sep 1: $4.64 Sep 2: $4.64 Sep 4: $4.64 Sep 5: $4.64 Sep 6: $4.64 Jun 15 Sep 6

Budget

One cost-efficient workhorse per major provider

$0.74/Mtok

+4.9% since start

$0.70 $0.71 $0.73 $0.74 Jun 15: $0.70 Jun 16: $0.70 Jun 17: $0.70 Jun 19: $0.70 Jun 20: $0.70 Jun 21: $0.70 Jun 22: $0.70 Jun 23: $0.70 Jun 24: $0.70 Jun 25: $0.70 Jun 26: $0.70 Jun 27: $0.70 Jun 28: $0.70 Jun 29: $0.70 Jun 30: $0.70 Jul 1: $0.70 Jul 2: $0.70 Jul 3: $0.70 Jul 4: $0.70 Jul 5: $0.70 Jul 6: $0.70 Jul 7: $0.70 Jul 8: $0.70 Jul 9: $0.70 Jul 10: $0.70 Jul 11: $0.70 Jul 12: $0.70 Jul 13: $0.70 Jul 14: $0.70 Jul 15: $0.70 Jul 16: $0.70 Jul 17: $0.70 Jul 18: $0.70 Jul 19: $0.70 Jul 20: $0.70 Jul 21: $0.70 Jul 22: $0.70 Jul 23: $0.70 Jul 24: $0.70 Jul 25: $0.70 Jul 26: $0.70 Jul 27: $0.70 Jul 28: $0.70 Jul 29: $0.70 Jul 30: $0.70 Jul 31: $0.70 Aug 1: $0.70 Aug 2: $0.70 Aug 3: $0.70 Aug 4: $0.70 Aug 6: $0.70 Aug 8: $0.70 Aug 9: $0.70 Aug 10: $0.70 Aug 11: $0.70 Aug 12: $0.70 Aug 13: $0.70 Aug 14: $0.70 Aug 15: $0.70 Aug 16: $0.70 Aug 17: $0.70 Aug 18: $0.70 Aug 19: $0.70 Aug 21: $0.74 Aug 22: $0.74 Aug 23: $0.74 Aug 24: $0.74 Aug 25: $0.74 Aug 26: $0.74 Aug 27: $0.74 Aug 28: $0.74 Aug 29: $0.74 Aug 30: $0.74 Aug 31: $0.74 Sep 1: $0.74 Sep 2: $0.74 Sep 4: $0.74 Sep 5: $0.74 Sep 6: $0.74 Jun 15 Sep 6

Price war

Biggest recent price moves

When a provider changes a price, it shows up here. Green is a cut, red is a hike (blended $/Mtok).

xai/grok-4-fast-reasoning

Xai · $0.28 → $1.56

+468.2%

xai/grok-4-fast-non-reasoning

Xai · $0.28 → $1.56

+468.2%

xai/grok-4-1-fast

Xai · $0.28 → $1.56

+468.2%

xai/grok-4-1-fast-reasoning

Xai · $0.28 → $1.56

+468.2%

xai/grok-4-1-fast-reasoning-latest

Xai · $0.28 → $1.56

+468.2%

xai/grok-4-1-fast-non-reasoning

Xai · $0.28 → $1.56

+468.2%

Intelligence & speed

Not just cheap, capable

Independent quality and speed benchmarks, so cost reads next to capability. Smartest, fastest, and the best capability per dollar.

🧠 Smartest

Intelligence Index

  1. 1 Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback) Anthropic 56.8
  2. 2 GPT-6 Astra (max) OpenAI 54.7
  3. 3 GPT-6 Astra (xhigh) OpenAI 54.3
  4. 4 Claude Opus 5 (Adaptive Reasoning, Xhigh Effort) Anthropic 53.4
  5. 5 GPT-5.5 (medium) OpenAI 50.4
  6. 6 Claude Fable 5.1 (Adaptive Reasoning, Medium Effort, Default Fallback) Anthropic 49.9

⚡ Fastest

Output speed

  1. 1 Mercury 2 Inception 821 tok/s
  2. 2 Step 3.7 Flash StepFun 413 tok/s
  3. 3 Gemini 3.5 Flash-Lite Google 371 tok/s
  4. 4 LFM2.5-VL-1.6B Liquid AI 358 tok/s
  5. 5 gpt-oss-120b (low) OpenAI 346 tok/s
  6. 6 HyperNova 60B 2605 (high, based on gpt-oss-120b) Multiverse Computing 345 tok/s

💎 Best value

Intelligence per $/Mtok

  1. 1 Gemma 4 E4B (Non-reasoning) Google 217.5 pts/$
  2. 2 DeepSeek V4 Flash (Reasoning, High Effort) DeepSeek 213.7 pts/$
  3. 3 Granite 4.2 3B IBM 203.8 pts/$
  4. 4 GLM-5.3-Flash Z AI 194.5 pts/$
  5. 5 HyperNova 60B 2605 (high, based on gpt-oss-120b) Multiverse Computing 180 pts/$
  6. 6 Nemotron 3.5 Lightning NVIDIA 172.6 pts/$

Quality & speed data from Artificial Analysis.

Human preference

What people actually prefer

The LMArena (Chatbot Arena) overall leaderboard, decided by millions of blind head-to-head votes. Snapshot Aug 27, 2026.

  • 1 claude-opus-5-max Anthropic 1505 15.4K votes
  • 2 claude-opus-5-high Anthropic 1504 31.6K votes
  • 3 claude-opus-4-6-high Anthropic 1503 72.1K votes
  • 4 claude-opus-4-6 Anthropic 1498 76.1K votes
  • 5 claude-fable-5 Anthropic 1495 25.8K votes
  • 6 gemini-3.7-flash-high Google 1490 5.7K votes
  • 7 claude-opus-4-7-high Anthropic 1490 60.1K votes
  • 8 muse-spark-1.2 (xHigh) Meta 1488 3.2K votes
  • 9 gemini-3.5-flash-high Google 1483 32.2K votes
  • 10 claude-opus-4-7 Anthropic 1483 61.3K votes

Data: LMArena leaderboard, CC BY 4.0.

Real usage

What the community really spends

Anonymized, aggregate, opt-in. Numbers from real engineers, not list prices.

Tracked spend

$1,024.61

Tokens

117.7B

Tokens / $

114.9M

Contributors

6

Community dashboard →

Daily community spend

$0.00 $15.2 $30.5 $45.7 May 22: $6.98 May 23: $1.90 Jun 2: $12.9 Jun 3: $28.6 Jun 4: $11.5 Jun 10: $0.88 Jun 11: $5.14 Jun 12: $13.2 Jun 15: $18.1 Jun 16: $5.31 Jun 17: $0.74 Jun 18: $6.15 Jun 19: $28.9 Jun 20: $36.9 Jun 21: $0.15 Jun 22: $11.9 Jun 23: $7.57 Jun 25: $0.92 Jun 28: $0.43 Jun 30: $2.43 Jul 1: $1.35 Jul 2: $16.9 Jul 3: $18.2 Jul 5: $17.6 Jul 7: $20.9 Jul 8: $12.8 Jul 9: $4.87 Jul 11: $29.5 Jul 12: $20.3 Jul 13: $11.3 Jul 15: $8.49 Jul 16: $33.2 Jul 17: $40.8 Jul 18: $15.9 Jul 19: $24.3 Jul 21: $5.59 Jul 22: $17.8 Jul 23: $13.1 Jul 24: $21.6 Jul 26: $7.99 Jul 27: $11.1 Jul 28: $12.2 Jul 29: $9.60 Jul 30: $22.6 Jul 31: $36.3 Aug 1: $0.69 Aug 2: $27.6 Aug 3: $13.5 Aug 4: $33.3 Aug 5: $29.4 Aug 6: $26.3 Aug 7: $26.5 Aug 10: $19.4 Aug 11: $8.26 Aug 12: $13.8 Aug 13: $12.1 Aug 14: $14.4 Aug 15: $6.65 Aug 18: $13.1 Aug 19: $13.2 Aug 20: $13.8 Aug 21: $5.86 Aug 24: $6.87 Aug 25: $8.47 Aug 26: $4.03 Aug 27: $8.58 Aug 28: $6.30 Aug 29: $0.43 Aug 30: $15.5 Aug 31: $39.0 Sep 3: $9.39 Sep 4: $9.60 Sep 5: $3.96 May 22 Sep 5

By provider

  • Anthropic 100%

By platform

  • Claude Code 100%

By use case

  • Coding 100%

Value

Cheapest & priciest right now

Blended $/Mtok (3:1 input:output) across text models with public prices.

Cheapest

  • fireworks_ai/accounts/fireworks/models/flux-1-dev-controlnet-union Fireworks Ai $0/Mtok
  • fireworks-ai-embedding-up-to-150m Fireworks Ai Embedding Models $0.01/Mtok
  • fireworks-ai-embedding-150m-to-350m Fireworks Ai Embedding Models $0.01/Mtok
  • nscale/Qwen/Qwen2.5-Coder-3B-Instruct Nscale $0.02/Mtok
  • nscale/Qwen/Qwen2.5-Coder-7B-Instruct Nscale $0.02/Mtok
  • nebius/Qwen/Qwen2.5-Coder-7B Nebius $0.02/Mtok

Priciest

  • wandb/zai-org/GLM-4.5 Wandb $91,250/Mtok
  • wandb/microsoft/Phi-4-mini-instruct Wandb $14,750/Mtok
  • azure_ai/jais-30b-chat Azure $4,827.5/Mtok
  • watsonx/core42/jais-13b-chat Watsonx $875/Mtok
  • openrouter/openai/o1-pro Openrouter $262.5/Mtok
  • o1-pro-2025-03-19 Openai $262.5/Mtok

Explore

Every model, your way

Filter by provider, search, and sort the entire live price catalog. Pulled straight from the open Data API.

Model Provider Input Output Blended

Source: GET /api/v1/data/models · CC BY 4.0

The cheatsheet

Which AI for which job

No single model wins everything. The top pick for each job, ranked from real data.

🧠

Reasoning

Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback)

💻

Coding

Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback)

🤖

Agents

Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback)

💬

Chat & writing

claude-opus-5-max

Speed

Step 3.7 Flash

📚

Long context

Llama 4 Scout 17b 128e Instruct Maas

💎

Best value

Gemma 4 E4B (Non-reasoning)

🪙

Budget

Llama 3.1 8b

Full cheatsheet →

The price sheet lies

List price isn't what you pay

Ring-2.6-1T is listed 92% cheaper than GPT-5.5 (Non-reasoning) — yet costs 138% more to actually run. Each model by list price (left) vs the real bill to run the Intelligence Index (right).

LIST PRICE $ per million tokens ACTUAL COST $ to run the Intelligence Index Granite 4.2 3B$0.05 Granite 4.2 3B — list $0.05/Mtok GR gpt-oss-120b (high)$0.26 gpt-oss-120b (high) — list $0.26/Mtok GP GPT-5.4 nano (xhigh)$0.46 GPT-5.4 nano (xhigh) — list $0.46/Mtok GT MiniMax-M3$0.53 MiniMax-M3 — list $0.53/Mtok MI Gemini 3.1 Flash-Lite$0.56 Gemini 3.1 Flash-Lite — list $0.56/Mtok GE Ring-2.6-1T$0.85 Ring-2.6-1T — list $0.85/Mtok RI Inkling (xhigh)$1.76 Inkling (xhigh) — list $1.76/Mtok IN Muse Spark 1.2 (xhigh)$2.00 Muse Spark 1.2 (xhigh) — list $2.00/Mtok MU GPT-5.1 (high)$3.44 GPT-5.1 (high) — list $3.44/Mtok G5 Kimi K3 (low)$6.00 Kimi K3 (low) — list $6.00/Mtok KI Claude Opus 5 (Adaptive…$10.00 Claude Opus 5 (Adaptive Reasoning, Low Effort) — list $10.00/Mtok CL GPT-5.5 (Non-reasoning)$11.25 GPT-5.5 (Non-reasoning) — list $11.25/Mtok G2 GPT-6 Astra (Non-reason…$20.00 GPT-6 Astra (Non-reasoning) — list $20.00/Mtok G6 Claude Fable 5.1 (Adapt…$20.00 Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback) — list $20.00/Mtok CA $14Granite 4.2 3B Granite 4.2 3B — actually $14 to run the Intelligence Index ($0.0055/task) GR $96gpt-oss-120b (high) gpt-oss-120b (high) — actually $96 to run the Intelligence Index ($0.0607/task) GP $267GPT-5.4 nano (xhigh) GPT-5.4 nano (xhigh) — actually $267 to run the Intelligence Index ($0.1228/task) GT $293MiniMax-M3 MiniMax-M3 — actually $293 to run the Intelligence Index ($0.2294/task) MI $94Gemini 3.1 Flash-Lite Gemini 3.1 Flash-Lite — actually $94 to run the Intelligence Index ($0.0416/task) GE $459Ring-2.6-1T Ring-2.6-1T — actually $459 to run the Intelligence Index ($0.3448/task) RI $826Inkling (xhigh) Inkling (xhigh) — actually $826 to run the Intelligence Index ($0.4379/task) IN $822Muse Spark 1.2 (xhigh) Muse Spark 1.2 (xhigh) — actually $822 to run the Intelligence Index ($0.5453/task) MU $782GPT-5.1 (high) GPT-5.1 (high) — actually $782 to run the Intelligence Index ($0.3040/task) G5 $283Kimi K3 (low) Kimi K3 (low) — actually $283 to run the Intelligence Index ($0.2419/task) KI $811Claude Opus 5 (Adapti… Claude Opus 5 (Adaptive Reasoning, Low Effort) — actually $811 to run the Intelligence Index ($0.6358/task) CL $193GPT-5.5 (Non-reasonin… GPT-5.5 (Non-reasoning) — actually $193 to run the Intelligence Index ($0.1515/task) G2 $1,631GPT-6 Astra (Non-reas… GPT-6 Astra (Non-reasoning) — actually $1,631 to run the Intelligence Index ($1.4233/task) G6 $10,816Claude Fable 5.1 (Ada… Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback) — actually $10,816 to run the Intelligence Index ($6.1169/task) CA
Pricier than it looks Secret bargain
Full breakdown →

Run cost = cost to run the full Intelligence Index, from Artificial Analysis.

Citation

Use this in your work

Quote any figure on this page. It is all open data. Here is a ready-made citation, and the methodology behind every number.

Copy a citation

Free to use and cite under CC BY 4.0. See how this is measured.

APA

Champlin Enterprises. (2026). The State of AI (MyTokenTracker) [Data set]. MyTokenTracker. Retrieved September 7, 2026, from https://mytokentracker.io/state-of-ai

BibTeX
@misc{mytokentracker-state-of-ai,
  title        = {The State of AI (MyTokenTracker)},
  author       = {{Champlin Enterprises}},
  year         = {2026},
  howpublished = {MyTokenTracker, \url{https://mytokentracker.io/state-of-ai}},
  note         = {Accessed September 7, 2026. Licensed CC BY 4.0.},
  url          = {https://mytokentracker.io/state-of-ai}
}

Need a fixed point in time? Every day’s data is permanently archived in the open-data repository, so you can cite a specific date by linking that day’s committed file.

Free weekly digest

Get this in your inbox, weekly

The State of AI moves every day. Once a week we send the headline: the AI Cost Index, the biggest price moves, and what it means. Free, no account.

No spam, no account. One click to leave.

Measure it. Don't guess.

Every number on this page came from people who track instead of guess. One line to install, automatic capture, free forever. Add your usage and the whole picture gets sharper.

Our wiggly friend is an original MyTokenTracker mascot, here purely for fun. It is not affiliated with, endorsed by, or representing Anthropic, OpenAI, Google, or any model provider. Product and model names (Claude, GPT, Gemini, and others) are trademarks of their respective owners and appear here only to report public pricing and usage.