Context
262K
Max output
n/a
Serving providers
1
Cheapest input
$0.08 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: Artificial Analysis
Speed & latency
Output speed
291 t/s
tokens / second
Time to first token
0.84s
latency
Headline indices
26.8
of 100
13.8
of 100
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
CritPt0%default
EvalComputeProxy238.1default
GDPval823.6default
GPQA Diamond74%default
Humanity's Last Exam11%default
Long-Context Reasoning55%default
Mlcr Overall0.0default
Omniscience-17.7default
Omniscience Accuracy0.1default
Omniscience Non Hallucination0.6default
SciCode32%default
TerminalBench v2.124%default
τ-bench Banking9%default

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
OpenRouterCheapest$0.08$0.201
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$0.08$0.20$0.04nvidia/nemotron-3.5-lightning

Price history

input + output $/1M since we started tracking
Input Output
Input down 20% since first tracked
$0.000$0.100$0.200$0.300Aug 12Aug 17Aug 28Aug 29$0.200$0.080

Cost calculator

estimate your monthly spend on this model
$32
estimated / month

Model IDs

copy the exact identifier for your platform
nvidia/nemotron-3.5-lightning

Frequently asked questions

Nemotron 3.5 Lightning pricing, context and availability

How much does Nemotron 3.5 Lightning cost?

Nemotron 3.5 Lightning costs $0.08 per 1M input tokens and $0.20 per 1M output tokens at its cheapest provider via OpenRouter.

What is the context window of Nemotron 3.5 Lightning?

Nemotron 3.5 Lightning accepts up to 262K tokens of context.

What can Nemotron 3.5 Lightning do?

Nemotron 3.5 Lightning supports tool use and reasoning.

Other NVIDIA models

compare pricing across the NVIDIA lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI