This model was retired on 2026-08-16. The pricing below is the last-known rate, kept for migration reference.

Llama 3.1 8B InstantDeprecated

Meta
Tool use
Context
131K
Max output
131K
Serving providers
1
Cheapest input
$0.05 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Pricing by serving provider

last-known · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
Groq$0.05$0.081

Price history

input + output $/1M since we started tracking
Input Output
Input down 92% since first tracked
$0.000$0.200$0.400$0.600$0.800Jul 25Oct 8Jun 22$0.080$0.050

Cost calculator

estimate your monthly spend on this model
$16
estimated / month

Model IDs

copy the exact identifier for your platform
groq/llama-3.1-8b-instant

Frequently asked questions

Llama 3.1 8B Instant pricing, context and availability

Is Llama 3.1 8B Instant still available?

Llama 3.1 8B Instant was retired on 2026-08-16. The pricing on this page is the last-known rate, kept for migration reference.

How much did Llama 3.1 8B Instant cost?

Llama 3.1 8B Instant's last-known pricing, before it was retired on 2026-08-16, was $0.05 per 1M input tokens and $0.08 per 1M output tokens.

What was the context window of Llama 3.1 8B Instant?

Llama 3.1 8B Instant had a 131K token context window and could return up to 131K output tokens.

What could Llama 3.1 8B Instant do?

Llama 3.1 8B Instant supported tool use.

Other Meta models

compare pricing across the Meta lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI