Context
512K
Max output
n/a
Serving providers
1
Cheapest input
$0.60 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
OpenRouterCheapest$0.60$3.601
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$0.60$3.60$0.20nvidia/nemotron-3-ultra-550b-a55b:batch

Price history

input + output $/1M since we started tracking
Input Output
$0.000$1.00$2.00$3.00$4.00Aug 7Aug 19$3.60$0.600

Cost calculator

estimate your monthly spend on this model
$408
estimated / month

Model IDs

copy the exact identifier for your platform
nvidia/nemotron-3-ultra-550b-a55b:batch

Frequently asked questions

Nemotron 3 Ultra (batch) pricing, context and availability

How much does Nemotron 3 Ultra (batch) cost?

Nemotron 3 Ultra (batch) costs $0.60 per 1M input tokens and $3.60 per 1M output tokens at its cheapest provider via OpenRouter.

What is the context window of Nemotron 3 Ultra (batch)?

Nemotron 3 Ultra (batch) accepts up to 512K tokens of context.

Other NVIDIA models

compare pricing across the NVIDIA lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI