This model was retired on 2026-03-09. The pricing below is the last-known rate, kept for migration reference.

Llama 4 Maverick 17B 128e InstructDeprecated

Meta
VisionTool useStructured output
Context
131K
Max output
131K
Serving providers
2
Cheapest input
$0.20 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Pricing by serving provider

last-known · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
Groq$0.20$0.601
SambaNova$0.63$1.801

Price history

input + output $/1M since we started tracking
Input Output
$0.000$0.200$0.400$0.600May 14May 19Jun 22$0.600$0.200

Cost calculator

estimate your monthly spend on this model
$88
estimated / month

Model IDs

copy the exact identifier for your platform
groq/meta-llama/llama-4-maverick-17b-128e-instructsambanova/Llama-4-Maverick-17B-128E-Instruct

Frequently asked questions

Llama 4 Maverick 17B 128e Instruct pricing, context and availability

Is Llama 4 Maverick 17B 128e Instruct still available?

Llama 4 Maverick 17B 128e Instruct was retired on 2026-03-09. The pricing on this page is the last-known rate, kept for migration reference.

How much did Llama 4 Maverick 17B 128e Instruct cost?

Llama 4 Maverick 17B 128e Instruct's last-known pricing, before it was retired on 2026-03-09, was $0.20 per 1M input tokens and $0.60 per 1M output tokens.

What was the context window of Llama 4 Maverick 17B 128e Instruct?

Llama 4 Maverick 17B 128e Instruct had a 131K token context window and could return up to 131K output tokens.

What could Llama 4 Maverick 17B 128e Instruct do?

Llama 4 Maverick 17B 128e Instruct supported image input (vision), tool use and structured output.

Other Meta models

compare pricing across the Meta lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI