Llama 4 Scout

MetaReleasedJan 1, 2025KnowledgeJan 2025
Context
131K
Max output
8K
Serving providers
2
Cheapest input
$0.10 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: Artificial Analysis
Speed & latency
Output speed
131 t/s
tokens / second
Time to first token
0.80s
latency
Headline indices
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
CritPt0%default
EvalComputeProxy81.4default
GDPval108.6default
GPQA Diamond59%default
Humanity's Last Exam4%default
IFBench40%default
LiveCodeBench30%default
Long-Context Reasoning30%default
MMLU-Pro75%default
MMMU-Pro53%default
Omniscience-52.1default
Omniscience Accuracy0.2default
Omniscience Non Hallucination0.2default
SciCode17%default
TerminalBench Hard2%default
TerminalBench v2.14%default
τ-bench Banking3%default
τ²-bench15%default

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand

Input pricing runs from $0.10 to $0.11 per 1M tokens across 2 providers, so the dearest route costs 10% more than the cheapest for the same model.

Serving providerInput /1MOutput /1MEndpoints
Vercel AI GatewayCheapest$0.10$0.301
ComponentUnitStandardBatchCached
Text input/1M tok$0.1000
Text output/1M tok$0.3000
OpenRouter$0.11$0.341

Price history

input + output $/1M since we started tracking
Input Output
$0.000$0.100$0.200$0.300Jul 30Jun 22Jul 19Aug 26$0.300$0.100

Cost calculator

estimate your monthly spend on this model
$44
estimated / month

Model IDs

copy the exact identifier for your platform
vercel_ai_gateway/meta/llama-4-scoutmeta-llama/llama-4-scout

Frequently asked questions

Llama 4 Scout pricing, context and availability

How much does Llama 4 Scout cost?

Llama 4 Scout costs $0.10 per 1M input tokens and $0.30 per 1M output tokens at its cheapest provider via Vercel AI Gateway. Across 2 serving providers, input prices range from $0.10 to $0.11 per 1M tokens.

What is the context window of Llama 4 Scout?

Llama 4 Scout accepts up to 131K tokens of context and can return up to 8K output tokens.

Which providers serve Llama 4 Scout?

Llama 4 Scout is available from 2 serving providers, each with its own pricing and model ID. Vercel AI Gateway is currently the cheapest.

What can Llama 4 Scout do?

Llama 4 Scout supports image input (vision) and tool use.

Other Meta models

compare pricing across the Meta lineup

Popular comparisons

head-to-head pages featuring Llama 4 Scout
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI