baseten/thinkingmachines/inkling
Baseten textbaseten/thinkingmachines/inkling
Input
$1.00
per 1M input tokens
Output
$4.05
per 1M output tokens
Cache read
$0.1700
per 1M cached read tokens
Cache write
n/a
per 1M cache write tokens
Context window
1,048,576
Max output
32,768
Effective date
Sep 23, 2026
Estimate a workload
Enter token counts to see the cost at this model's rates.
Estimated cost:
$0.00
Price history
blended $ per 1M tokens (3:1)One price on record so far (Sep 23, 2026). This chart fills in as the price changes over time. We check daily.
| Effective | Input | Output | Blended |
|---|---|---|---|
| Sep 23, 2026 | $1.00 | $4.05 | $1.76 |
Other Baseten models
baseten/openai/gpt-oss-120b
· $0.5000/Mtok out
baseten/deepseek-ai/DeepSeek-V4-Flash-0731
· $0.2600/Mtok out
baseten/zai-org/GLM-5.3-Flash
· $0.5000/Mtok out
baseten/deepseek-ai/DeepSeek-V4.1-Flash
· $1.20/Mtok out
baseten/MiniMaxAI/MiniMax-M2.5
· $1.20/Mtok out
baseten/nvidia/Nemotron-120B-A12B
· $0.7500/Mtok out
Source: litellm. Confirm against the provider's official pricing before relying on these figures.