The price sheet lies

List price tells you almost nothing about what AI really costs.

Muse Spark 1.2 (xhigh) is listed 82% cheaper than GPT-5.5 (low). It costs 287% more to actually run.

The flip

Same models, two orderings

Each model ranked by list price on the left, and by what it actually costs to run Artificial Analysis's Intelligence Index on the right. Follow a line across — where it climbs, the sticker price lied.

LIST PRICE $ per million tokens ACTUAL COST $ to run the Intelligence Index Granite 4.2 3B$0.05 Granite 4.2 3B — list $0.05/Mtok GR Mistral Small 3.2$0.15 Mistral Small 3.2 — list $0.15/Mtok MI Mistral Small 4 (Reason…$0.26 Mistral Small 4 (Reasoning) — list $0.26/Mtok MS DeepSeek V3 (Dec '24)$0.46 DeepSeek V3 (Dec '24) — list $0.46/Mtok DE MiMo-V2.5-Pro$0.54 MiMo-V2.5-Pro — list $0.54/Mtok MM Mistral Large 3$0.75 Mistral Large 3 — list $0.75/Mtok MT Nemotron 3 Ultra 550B A…$1.18 Nemotron 3 Ultra 550B A55B (Reasoning) — list $1.18/Mtok NE Muse Spark 1.2 (xhigh)$2.00 Muse Spark 1.2 (xhigh) — list $2.00/Mtok MU Grok 4.6 (medium)$3.00 Grok 4.6 (medium) — list $3.00/Mtok GO Claude Sonnet 5 (Adapti…$4.00 Claude Sonnet 5 (Adaptive Reasoning, Low Effort) — list $4.00/Mtok CL Claude Sonnet 4.6 (Adap…$6.00 Claude Sonnet 4.6 (Adaptive Reasoning, Max Effort) — list $6.00/Mtok CA GPT-5.5 (low)$11.25 GPT-5.5 (low) — list $11.25/Mtok GP Claude Fable 5.1 (Adapt…$20.00 Claude Fable 5.1 (Adaptive Reasoning, Low Effort, Default Fallback) — list $20.00/Mtok CU Claude Fable 5.1 (Adapt…$20.00 Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback) — list $20.00/Mtok CD $15Granite 4.2 3B Granite 4.2 3B — actually $15 to run the Intelligence Index ($0.0060/task) GR $163Mistral Small 3.2 Mistral Small 3.2 — actually $163 to run the Intelligence Index ($0.1458/task) MI $79Mistral Small 4 (Reas… Mistral Small 4 (Reasoning) — actually $79 to run the Intelligence Index ($0.0451/task) MS $25DeepSeek V3 (Dec '24) DeepSeek V3 (Dec '24) — actually $25 to run the Intelligence Index ($0.0198/task) DE $138MiMo-V2.5-Pro MiMo-V2.5-Pro — actually $138 to run the Intelligence Index ($0.0541/task) MM $130Mistral Large 3 Mistral Large 3 — actually $130 to run the Intelligence Index ($0.1044/task) MT $443Nemotron 3 Ultra 550B… Nemotron 3 Ultra 550B A55B (Reasoning) — actually $443 to run the Intelligence Index ($0.2439/task) NE $1,385Muse Spark 1.2 (xhigh) Muse Spark 1.2 (xhigh) — actually $1,385 to run the Intelligence Index ($0.9747/task) MU $1,937Grok 4.6 (medium) Grok 4.6 (medium) — actually $1,937 to run the Intelligence Index ($1.4963/task) GO $653Claude Sonnet 5 (Adap… Claude Sonnet 5 (Adaptive Reasoning, Low Effort) — actually $653 to run the Intelligence Index ($0.5088/task) CL $3,356Claude Sonnet 4.6 (Ad… Claude Sonnet 4.6 (Adaptive Reasoning, Max Effort) — actually $3,356 to run the Intelligence Index ($1.1356/task) CA $358GPT-5.5 (low) GPT-5.5 (low) — actually $358 to run the Intelligence Index ($0.1930/task) GP $3,158Claude Fable 5.1 (Ada… Claude Fable 5.1 (Adaptive Reasoning, Low Effort, Default Fallback) — actually $3,158 to run the Intelligence Index ($2.3710/task) CU $13,129Claude Fable 5.1 (Ada… Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback) — actually $13,129 to run the Intelligence Index ($7.6297/task) CD
Cheaper on paper, pricier to run Looks pricey, secretly a bargain Every other measured model

14 models with both a list price and a measured run cost · updated 9 hours ago

Why the sticker lies

You don't buy tokens. You buy answers.

A price sheet quotes dollars per million tokens. But a task isn't a fixed number of tokens — and that's where the bill hides.

List price is per token

The number on the pricing page is $ per million input/output tokens. It says nothing about how many tokens a model will spend to finish your task.

Reasoning models are verbose

A "thinking" model can emit 10–50× more tokens chewing through the same problem. Cheap per token × a mountain of tokens = an expensive answer.

The real unit is the task

Run the same fixed benchmark across models and the cheap-looking ones often cost the most. Rank by the finished job, not the sticker.

Receipts

Where list price misleads the most

Sorted by how far a model moves when you re-rank by real cost. ▲ means it's pricier than its sticker suggests; ▼ means it's a quiet bargain.

Model List $/Mtok Actual run cost Re-rank
Mistral Small 3.2 Mistral $0.15 $163 ▲ 4 pricier
GPT-5.5 (low) OpenAI $11.25 $358 ▼ 5 cheaper
Muse Spark 1.2 (xhigh) Meta $2.00 $1,385 ▲ 2 pricier
Grok 4.6 (medium) SpaceXAI $3.00 $1,937 ▲ 2 pricier
Claude Sonnet 4.6 (Adaptive Reasoning, Max Effort) Anthropic $6.00 $3,356 ▲ 2 pricier
Mistral Large 3 Mistral $0.75 $130 ▼ 2 cheaper
DeepSeek V3 (Dec '24) DeepSeek $0.46 $25 ▼ 2 cheaper
Nemotron 3 Ultra 550B A55B (Reasoning) NVIDIA $1.18 $443 ▲ 1 pricier
Claude Sonnet 5 (Adaptive Reasoning, Low Effort) Anthropic $4.00 $653 ▼ 1 cheaper
Claude Fable 5.1 (Adaptive Reasoning, Low Effort, Default Fallback) Anthropic $20.00 $3,158 ▼ 1 cheaper

List price is the blended 3:1 input:output rate. "Actual run cost" is what it costs to run the full Artificial Analysis Intelligence Index, from Artificial Analysis — a fixed task suite, so the only variable is how each model behaves.

Citation

Use this in your work

Every figure here is open data. Here's a ready-made citation, plus the methodology behind the numbers and the full State of AI.

Copy a citation

Free to use and cite under CC BY 4.0. See how this is measured.

APA

Champlin Enterprises. (2026). The AI Price Sheet Lies (MyTokenTracker) [Data set]. MyTokenTracker. Retrieved September 9, 2026, from https://mytokentracker.io/price-vs-cost

BibTeX
@misc{mytokentracker-price-vs-cost,
  title        = {The AI Price Sheet Lies (MyTokenTracker)},
  author       = {{Champlin Enterprises}},
  year         = {2026},
  howpublished = {MyTokenTracker, \url{https://mytokentracker.io/price-vs-cost}},
  note         = {Accessed September 9, 2026. Licensed CC BY 4.0.},
  url          = {https://mytokentracker.io/price-vs-cost}
}

Need a fixed point in time? Every day’s data is permanently archived in the open-data repository, so you can cite a specific date by linking that day’s committed file.

Free weekly digest

Stop guessing what AI costs

We track the real, measured cost of every model — not the sticker price. One line to install, free forever. Get the weekly headline in your inbox.

No spam, no account. One click to leave.