This model was retired on 2026-06-13. The pricing below is the last-known rate, kept for migration reference.
Context
131K
Max output
131K
Serving providers
5
Cheapest input
$0.05 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.
Benchmarks
independent evaluations · source: Artificial AnalysisSpeed & latency
Output speed
30 t/s
tokens / second
Time to first token
1.22s
latency
Headline indices
Intelligence Index
3.0
of 100
Individual benchmarks · click a row for effort variants
| Benchmark | Best score | Effort | Cost / task | |
|---|---|---|---|---|
| CritPt | 0% | default | — | |
| EvalComputeProxy | 3.2 | default | — | |
| GPQA Diamond | 22% | default | — | |
| Humanity's Last Exam | 6% | default | — | |
| IFBench | 30% | default | — | |
| LiveCodeBench | 11% | default | — | |
| Long-Context Reasoning | 16% | default | — | |
| MMLU-Pro | 46% | default | — | |
| MMMU-Pro | 29% | default | — | |
| Omniscience | -62.6 | default | — | |
| Omniscience Accuracy | 0.1 | default | — | |
| Omniscience Non Hallucination | 0.2 | default | — | |
| SciCode | 11% | default | — | |
| TerminalBench Hard | 1% | default | — | |
| τ²-bench | 15% | default | — |
Pricing by serving provider
last-known · per 1M tokens · click a provider to expand| Serving provider | Input /1M | Output /1M | Endpoints | ||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Cloudflare | $0.05 | $0.68 | 1 | ||||||||||||||||
| |||||||||||||||||||
| DeepInfra | $0.05 | $0.05 | 1 | ||||||||||||||||
| |||||||||||||||||||
| Watsonx | $0.35 | $0.35 | 1 | ||||||||||||||||
| |||||||||||||||||||
| Azure | $0.37 | $0.37 | 1 | ||||||||||||||||
| |||||||||||||||||||
| OCI | $2.00 | $2.00 | 1 | ||||||||||||||||
| |||||||||||||||||||
Price history
input + output $/1M since we started tracking Input Output
Input down 87% since first tracked
Cost calculator
estimate your monthly spend on this model$64
estimated / month
Model IDs
copy the exact identifier for your platformcloudflare/@cf/meta/llama-3.2-11b-vision-instructdeepinfra/meta-llama/Llama-3.2-11B-Vision-Instructwatsonx/meta-llama/llama-3-2-11b-vision-instructazure_ai/Llama-3.2-11B-Vision-Instructoci/meta.llama-3.2-11b-vision-instruct
Frequently asked questions
Llama 3.2 11B Vision Instruct pricing, context and availabilityIs Llama 3.2 11B Vision Instruct still available?
Llama 3.2 11B Vision Instruct was retired on 2026-06-13. The pricing on this page is the last-known rate, kept for migration reference.
How much did Llama 3.2 11B Vision Instruct cost?
Llama 3.2 11B Vision Instruct's last-known pricing, before it was retired on 2026-06-13, was $0.05 per 1M input tokens and $0.05 per 1M output tokens.
What was the context window of Llama 3.2 11B Vision Instruct?
Llama 3.2 11B Vision Instruct had a 131K token context window and could return up to 131K output tokens.
What could Llama 3.2 11B Vision Instruct do?
Llama 3.2 11B Vision Instruct supported image input (vision) and tool use.
Other Meta models
compare pricing across the Meta lineupLlama 3.3 70B Instruct$0.12/1M · 13 providersLlama 4 Scout 17B 16e Instruct$0.05/1M · 10 providersLlama 3.1 8B Instruct$0.02/1M · 8 providersMeta Llama 3.1 8B Instruct$0.02/1M · 6 providersLlama 3.2 3B Instruct$0.02/1M · 6 providersLlama 4 Maverick 17B 128e Instruct FP8$0.05/1M · 6 providersMeta Llama 3.1 70B Instruct$0.12/1M · 5 providersLlama Guard 3 8B$0.02/1M · 4 providers
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily
Every weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI