Llama 3.3 Nemotron Super 49B V1 5
Context
131K
Max output
131K
Serving providers
1
Cheapest input
$0.10 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.
Pricing by serving provider
standard tier · per 1M tokens · click a provider to expand| Serving provider | Input /1M | Output /1M | Endpoints | ||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| DeepInfraCheapest | $0.10 | $0.40 | 1 | ||||||||||||||||
| |||||||||||||||||||
Price history
input + output $/1M since we started tracking Input Output
Cost calculator
estimate your monthly spend on this model$52
estimated / month
Model IDs
copy the exact identifier for your platformdeepinfra/nvidia/Llama-3.3-Nemotron-Super-49B-v1.5
Frequently asked questions
Llama 3.3 Nemotron Super 49B V1 5 pricing, context and availabilityHow much does Llama 3.3 Nemotron Super 49B V1 5 cost?
Llama 3.3 Nemotron Super 49B V1 5 costs $0.10 per 1M input tokens and $0.40 per 1M output tokens at its cheapest provider via DeepInfra.
What is the context window of Llama 3.3 Nemotron Super 49B V1 5?
Llama 3.3 Nemotron Super 49B V1 5 accepts up to 131K tokens of context and can return up to 131K output tokens.
What can Llama 3.3 Nemotron Super 49B V1 5 do?
Llama 3.3 Nemotron Super 49B V1 5 supports tool use.
Other NVIDIA models
compare pricing across the NVIDIA lineupNvidia Nemotron Nano 9B$0.04/1M · 3 providersNemotron 3 Nano 30B A3B$0.05/1M · 3 providersNvidia Nemotron Nano 12B$0.20/1M · 2 providersNvidia Nemotron 3 Ultra 550B A55b$0.50/1M · 2 providersNemotron 3 Ultra$0.50/1M · 2 providersNemotron 3 Nano 30B A3B (free)$0.00/1M · 1 providerNemotron 3 Nano Omni (free)$0.00/1M · 1 providerNemotron 3 Super (free)$0.00/1M · 1 provider
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily
Every weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI