wandb/nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B
Wandb textwandb/nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B
Input
$0.7500
per 1M input tokens
Output
$2.75
per 1M output tokens
Cache read
$0.1500
per 1M cached read tokens
Cache write
n/a
per 1M cache write tokens
Context window
262,144
Max output
n/a
Effective date
Aug 28, 2026
Estimate a workload
Enter token counts to see the cost at this model's rates.
Estimated cost:
$0.00
Price history
blended $ per 1M tokens (3:1)One price on record so far (Aug 28, 2026). This chart fills in as the price changes over time. We check daily.
| Effective | Input | Output | Blended |
|---|---|---|---|
| Aug 28, 2026 | $0.75 | $2.75 | $1.25 |
Other Wandb models
wandb/openai/gpt-oss-120b
· $0.1700/Mtok out
wandb/openai/gpt-oss-20b
· $0.1300/Mtok out
wandb/ibm-granite/granite-4.1-8b
· $0.1000/Mtok out
wandb/JetBrains/Mellum2-12B-A2.5B-Instruct
· $0.1000/Mtok out
wandb/OpenPipe/Qwen3-14B-Instruct
· $0.2200/Mtok out
wandb/Qwen/Qwen3-235B-A22B-Instruct-2507
· $0.1000/Mtok out
Source: litellm. Confirm against the provider's official pricing before relying on these figures.