This model was retired on 2026-08-16. The pricing below is the last-known rate, kept for migration reference.

Llama 3.3 70B VersatileDeprecated

Meta
Tool use
Context
131K
Max output
33K
Serving providers
1
Cheapest input
$0.59 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Pricing by serving provider

last-known · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
Groq$0.59$0.791

Price history

input + output $/1M since we started tracking
Input Output
$0.000$0.200$0.400$0.600$0.800Dec 7Jun 22$0.790$0.590

Cost calculator

estimate your monthly spend on this model
$181
estimated / month

Model IDs

copy the exact identifier for your platform
groq/llama-3.3-70b-versatile

Frequently asked questions

Llama 3.3 70B Versatile pricing, context and availability

Is Llama 3.3 70B Versatile still available?

Llama 3.3 70B Versatile was retired on 2026-08-16. The pricing on this page is the last-known rate, kept for migration reference.

How much did Llama 3.3 70B Versatile cost?

Llama 3.3 70B Versatile's last-known pricing, before it was retired on 2026-08-16, was $0.59 per 1M input tokens and $0.79 per 1M output tokens.

What was the context window of Llama 3.3 70B Versatile?

Llama 3.3 70B Versatile had a 131K token context window and could return up to 33K output tokens.

What could Llama 3.3 70B Versatile do?

Llama 3.3 70B Versatile supported tool use.

Other Meta models

compare pricing across the Meta lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI