This model was retired on 2026-08-16. The pricing below is the last-known rate, kept for migration reference.
Context
131K
Max output
33K
Serving providers
1
Cheapest input
$0.59 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.
Pricing by serving provider
last-known · per 1M tokens · click a provider to expand| Serving provider | Input /1M | Output /1M | Endpoints | ||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Groq | $0.59 | $0.79 | 1 | ||||||||||||||||
| |||||||||||||||||||
Price history
input + output $/1M since we started tracking Input Output
Cost calculator
estimate your monthly spend on this model$181
estimated / month
Model IDs
copy the exact identifier for your platformgroq/llama-3.3-70b-versatile
Frequently asked questions
Llama 3.3 70B Versatile pricing, context and availabilityIs Llama 3.3 70B Versatile still available?
Llama 3.3 70B Versatile was retired on 2026-08-16. The pricing on this page is the last-known rate, kept for migration reference.
How much did Llama 3.3 70B Versatile cost?
Llama 3.3 70B Versatile's last-known pricing, before it was retired on 2026-08-16, was $0.59 per 1M input tokens and $0.79 per 1M output tokens.
What was the context window of Llama 3.3 70B Versatile?
Llama 3.3 70B Versatile had a 131K token context window and could return up to 33K output tokens.
What could Llama 3.3 70B Versatile do?
Llama 3.3 70B Versatile supported tool use.
Other Meta models
compare pricing across the Meta lineupLlama 3.3 70B Instruct$0.12/1M · 13 providersLlama 4 Scout 17B 16e Instruct$0.05/1M · 10 providersLlama 3.1 8B Instruct$0.02/1M · 8 providersMeta Llama 3.1 8B Instruct$0.02/1M · 6 providersLlama 3.2 3B Instruct$0.02/1M · 6 providersLlama 4 Maverick 17B 128e Instruct FP8$0.05/1M · 6 providersLlama 3.2 11B Vision Instruct$0.05/1M · 5 providersMeta Llama 3.1 70B Instruct$0.12/1M · 5 providers
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily
Every weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI