Context
262K
Max output
33K
Serving providers
1
Cheapest input
$0.05 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
Fireworks AICheapest$0.05$0.201
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$0.05$0.20$0.01fireworks_ai/accounts/fireworks/models/nemotron-lightning-3p5-30b-a3b

Price history

input + output $/1M since we started tracking
Input Output

Price history is accruing. We record a point each day a price changes; the full backfill lands shortly.

Cost calculator

estimate your monthly spend on this model
$26
estimated / month

Model IDs

copy the exact identifier for your platform
fireworks_ai/accounts/fireworks/models/nemotron-lightning-3p5-30b-a3b

Frequently asked questions

Nemotron Lightning 3p5 30B A3b pricing, context and availability

How much does Nemotron Lightning 3p5 30B A3b cost?

Nemotron Lightning 3p5 30B A3b costs $0.05 per 1M input tokens and $0.20 per 1M output tokens at its cheapest provider via Fireworks AI.

What is the context window of Nemotron Lightning 3p5 30B A3b?

Nemotron Lightning 3p5 30B A3b accepts up to 262K tokens of context and can return up to 33K output tokens.

What can Nemotron Lightning 3p5 30B A3b do?

Nemotron Lightning 3p5 30B A3b supports tool use, reasoning and structured output.

Other NVIDIA models

compare pricing across the NVIDIA lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI