This model was retired on 2026-03-05. The pricing below is the last-known rate, kept for migration reference.

Llama Guard 4 12BDeprecated

Meta
Context
1M
Max output
1M
Serving providers
4
Cheapest input
$0.18 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Pricing by serving provider

last-known · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
DeepInfra$0.18$0.181
OpenRouter$0.18$0.181
Groq$0.20$0.201
Together AI$0.20$0.201

Price history

input + output $/1M since we started tracking
Input Output
$0.000$0.050$0.100$0.150$0.200Jul 30Dec 16Jun 22Aug 26$0.180$0.180

Cost calculator

estimate your monthly spend on this model
$50
estimated / month

Model IDs

copy the exact identifier for your platform
deepinfra/meta-llama/Llama-Guard-4-12Bmeta-llama/llama-guard-4-12bgroq/meta-llama/llama-guard-4-12btogether_ai/meta-llama/Llama-Guard-4-12B

Frequently asked questions

Llama Guard 4 12B pricing, context and availability

Is Llama Guard 4 12B still available?

Llama Guard 4 12B was retired on 2026-03-05. The pricing on this page is the last-known rate, kept for migration reference.

How much did Llama Guard 4 12B cost?

Llama Guard 4 12B's last-known pricing, before it was retired on 2026-03-05, was $0.18 per 1M input tokens and $0.18 per 1M output tokens.

What was the context window of Llama Guard 4 12B?

Llama Guard 4 12B had a 1M token context window and could return up to 1M output tokens.

Other Meta models

compare pricing across the Meta lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI