Context
131K
Max output
128K
Serving providers
4
Cheapest input
$0.02 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: Artificial Analysis
Headline indices
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
CritPt0%default
EvalComputeProxy0.2default
GPQA Diamond20%default
Humanity's Last Exam5%default
IFBench23%default
LiveCodeBench2%default
Long-Context Reasoning6%default
MMLU-Pro20%default
Omniscience-54.7default
Omniscience Accuracy0.1default
Omniscience Non Hallucination0.3default
SciCode2%default
TerminalBench Hard0%default
τ²-bench0%default

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand

Input pricing runs from $0.02 to $0.10 per 1M tokens across 4 providers, so the dearest route costs 400% more than the cheapest for the same model.

Serving providerInput /1MOutput /1MEndpoints
NovitaCheapest$0.02$0.021
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$0.02$0.02novita/meta-llama/llama-3.2-1b-instruct
Cloudflare$0.03$0.201
OpenRouter$0.03$0.201
Watsonx$0.10$0.101

Price history

input + output $/1M since we started tracking
Input Output
$0.000$0.005$0.010Jul 30Oct 6Jun 24Aug 28$0.010$0.005

Cost calculator

estimate your monthly spend on this model
$6
estimated / month

Model IDs

copy the exact identifier for your platform
novita/meta-llama/llama-3.2-1b-instructcloudflare/@cf/meta/llama-3.2-1b-instructmeta-llama/llama-3.2-1b-instructwatsonx/meta-llama/llama-3-2-1b-instruct

Frequently asked questions

Llama 3.2 1B Instruct pricing, context and availability

How much does Llama 3.2 1B Instruct cost?

Llama 3.2 1B Instruct costs $0.02 per 1M input tokens and $0.02 per 1M output tokens at its cheapest provider via Novita. Across 4 serving providers, input prices range from $0.02 to $0.10 per 1M tokens.

What is the context window of Llama 3.2 1B Instruct?

Llama 3.2 1B Instruct accepts up to 131K tokens of context and can return up to 128K output tokens.

Which providers serve Llama 3.2 1B Instruct?

Llama 3.2 1B Instruct is available from 4 serving providers, each with its own pricing and model ID. Novita is currently the cheapest.

What can Llama 3.2 1B Instruct do?

Llama 3.2 1B Instruct supports tool use and structured output.

Other Meta models

compare pricing across the Meta lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI