Context
262K
Max output
n/a
Serving providers
1
Cheapest input
$0.08 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
DeepInfraCheapest$0.08$0.201
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$0.08$0.20$0.04deepinfra/nvidia/NVIDIA-Nemotron-3.5-Lightning

Price history

input + output $/1M since we started tracking
Input Output
$0.000$0.050$0.100$0.150$0.200Aug 14Aug 28$0.200$0.080

Cost calculator

estimate your monthly spend on this model
$32
estimated / month

Model IDs

copy the exact identifier for your platform
deepinfra/nvidia/NVIDIA-Nemotron-3.5-Lightning

Frequently asked questions

Nvidia Nemotron 3.5 Lightning pricing, context and availability

How much does Nvidia Nemotron 3.5 Lightning cost?

Nvidia Nemotron 3.5 Lightning costs $0.08 per 1M input tokens and $0.20 per 1M output tokens at its cheapest provider via DeepInfra.

What is the context window of Nvidia Nemotron 3.5 Lightning?

Nvidia Nemotron 3.5 Lightning accepts up to 262K tokens of context.

What can Nvidia Nemotron 3.5 Lightning do?

Nvidia Nemotron 3.5 Lightning supports tool use, reasoning and prompt caching.

Other NVIDIA models

compare pricing across the NVIDIA lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI