Benchmarks
independent evaluations · source: Artificial Analysis and Hugging Face leaderboards| Benchmark | Best score | Effort | Cost / task | |
|---|---|---|---|---|
| Aa Analyst Agent | 0.1 | default | — | |
| Automation Bench | 0.1 | default | — | |
| CritPt | 3% | default | — | |
| EvalComputeProxy | 2065.9 | default | — | |
| GDPval | 1163.0 | default | — | |
| GPQA Diamond | 87% | default | — | |
| Harvey Lab | 0.8 | default | — | |
| Humanity's Last Exam | 28% | default | — | |
| IFBench | 81% | default | — | |
| Long-Context Reasoning | 71% | default | — | |
| Mlcr Overall | 0.1 | default | — | |
| Omniscience | -0.4 | default | — | |
| Omniscience Accuracy | 0.2 | default | — | |
| Omniscience Non Hallucination | 0.7 | default | — | |
| SciCode | 40% | default | — | |
| SWE-bench Verified | 72% | default | — | |
| TerminalBench Hard | 36% | default | — | |
| TerminalBench v2.1 | 54% | default | — | |
| τ-bench Banking | 14% | default | — | |
| τ²-bench | 83% | default | — |
Pricing by serving provider
standard tier · per 1M tokens · click a provider to expandInput pricing runs from $0.50 to $0.75 per 1M tokens across 2 providers, so the dearest route costs 50% more than the cheapest for the same model.
| Serving provider | Input /1M | Output /1M | Endpoints | |||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| DeepInfraCheapest | $0.50 | $2.20 | 1 | |||||||||||||
| ||||||||||||||||
| Wandb | $0.75 | $2.75 | 1 | |||||||||||||
| ||||||||||||||||
Price history
input + output $/1M since we started trackingPrice history is accruing. We record a point each day a price changes; the full backfill lands shortly.
Cost calculator
estimate your monthly spend on this modelModel IDs
copy the exact identifier for your platformFrequently asked questions
Nvidia Nemotron 3 Ultra 550B A55b pricing, context and availabilityHow much does Nvidia Nemotron 3 Ultra 550B A55b cost?
Nvidia Nemotron 3 Ultra 550B A55b costs $0.50 per 1M input tokens and $2.20 per 1M output tokens at its cheapest provider via DeepInfra. Across 2 serving providers, input prices range from $0.50 to $0.75 per 1M tokens.
What is the context window of Nvidia Nemotron 3 Ultra 550B A55b?
Nvidia Nemotron 3 Ultra 550B A55b accepts up to 262K tokens of context.
Which providers serve Nvidia Nemotron 3 Ultra 550B A55b?
Nvidia Nemotron 3 Ultra 550B A55b is available from 2 serving providers, each with its own pricing and model ID. DeepInfra is currently the cheapest.
What can Nvidia Nemotron 3 Ultra 550B A55b do?
Nvidia Nemotron 3 Ultra 550B A55b supports image input (vision), tool use, reasoning, prompt caching and structured output.
Other NVIDIA models
compare pricing across the NVIDIA lineupEvery weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI