baseten/deepseek-ai/DeepSeek-V4-Flash-0731
Baseten textbaseten/deepseek-ai/DeepSeek-V4-Flash-0731
Input
$0.1300
per 1M input tokens
Output
$0.2600
per 1M output tokens
Cache read
$0.0280
per 1M cached read tokens
Cache write
n/a
per 1M cache write tokens
Context window
1,048,576
Max output
384,000
Effective date
Sep 23, 2026
Estimate a workload
Enter token counts to see the cost at this model's rates.
Estimated cost:
$0.00
Price history
blended $ per 1M tokens (3:1)One price on record so far (Sep 23, 2026). This chart fills in as the price changes over time. We check daily.
| Effective | Input | Output | Blended |
|---|---|---|---|
| Sep 23, 2026 | $0.13 | $0.26 | $0.16 |
Other Baseten models
baseten/openai/gpt-oss-120b
· $0.5000/Mtok out
baseten/zai-org/GLM-5.3-Flash
· $0.5000/Mtok out
baseten/deepseek-ai/DeepSeek-V4.1-Flash
· $1.20/Mtok out
baseten/MiniMaxAI/MiniMax-M2.5
· $1.20/Mtok out
baseten/nvidia/Nemotron-120B-A12B
· $0.7500/Mtok out
baseten/thinkingmachines/inkling-small
· $1.20/Mtok out
Source: litellm. Confirm against the provider's official pricing before relying on these figures.