Context
131K
Max output
131K
Serving providers
2
Cheapest input
$0.10 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand

Input pricing runs from $0.10 to $1.04 per 1M tokens across 2 providers, so the dearest route costs 940% more than the cheapest for the same model.

Serving providerInput /1MOutput /1MEndpoints
DeepInfraCheapest$0.10$0.321
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$0.10$0.32deepinfra/meta-llama/Llama-3.3-70B-Instruct-Turbo
Together AI$1.04$1.041

Price history

input + output $/1M since we started tracking
Input Output
Input down 89% since first tracked
$0.000$0.500$1.00Jan 20Sep 26Jun 22Aug 28$0.320$0.100

Cost calculator

estimate your monthly spend on this model
$46
estimated / month

Model IDs

copy the exact identifier for your platform
deepinfra/meta-llama/Llama-3.3-70B-Instruct-Turbotogether_ai/meta-llama/Llama-3.3-70B-Instruct-Turbo

Frequently asked questions

Llama 3.3 70B Instruct Turbo pricing, context and availability

How much does Llama 3.3 70B Instruct Turbo cost?

Llama 3.3 70B Instruct Turbo costs $0.10 per 1M input tokens and $0.32 per 1M output tokens at its cheapest provider via DeepInfra. Across 2 serving providers, input prices range from $0.10 to $1.04 per 1M tokens.

What is the context window of Llama 3.3 70B Instruct Turbo?

Llama 3.3 70B Instruct Turbo accepts up to 131K tokens of context and can return up to 131K output tokens.

Which providers serve Llama 3.3 70B Instruct Turbo?

Llama 3.3 70B Instruct Turbo is available from 2 serving providers, each with its own pricing and model ID. DeepInfra is currently the cheapest.

What can Llama 3.3 70B Instruct Turbo do?

Llama 3.3 70B Instruct Turbo supports tool use and structured output.

Other Meta models

compare pricing across the Meta lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI