deepinfra/meta-llama/Llama-4-Maverick-17B-128E-Instruct-FP8

Deepinfra text

deepinfra/meta-llama/Llama-4-Maverick-17B-128E-Instruct-FP8

Input

$0.2000

per 1M input tokens

Output

$0.8000

per 1M output tokens

Cache read

n/a

per 1M cached read tokens

Cache write

n/a

per 1M cache write tokens

Context window

1,048,576

Max output

1,048,576

Effective date

Aug 28, 2026

Estimate a workload

Enter token counts to see the cost at this model's rates.

Estimated cost: $0.00

Price history

blended $ per 1M tokens (3:1)
$0.25 $0.29 $0.32 $0.36 Jun 15, 2026: $0.26 Aug 28, 2026: $0.35 Jun 15, 2026 Aug 28, 2026
Effective Input Output Blended
Aug 28, 2026 $0.20 $0.80 $0.35
Jun 15, 2026 $0.15 $0.60 $0.26

Source: litellm. Confirm against the provider's official pricing before relying on these figures.