Context
n/a
Max output
n/a
Serving providers
0
Cheapest input
n/a /1M
Benchmarks
independent evaluations · source: Artificial AnalysisSpeed & latency
Output speed
100 t/s
tokens / second
Time to first token
0.91s
latency
Headline indices
Individual benchmarks · click a row for effort variants
| Benchmark | Best score | Effort | Cost / task | |
|---|---|---|---|---|
| CritPt | 3% | default | — | |
| EvalComputeProxy | 666.7 | default | — | |
| GDPval | 953 | default | — | |
| GPQA Diamond | 84% | default | — | |
| Humanity's Last Exam | 22% | default | — | |
| Long-Context Reasoning | 80% | default | — | |
| Mlcr Overall | 0.2 | default | — | |
| MMMU-Pro | 74% | default | — | |
| Omniscience | -32.9 | default | — | |
| Omniscience Accuracy | 0.3 | default | — | |
| Omniscience Non Hallucination | 0.2 | default | — | |
| SciCode | 44% | default | — | |
| TerminalBench v2.1 | 52% | default | — | |
| τ-bench Banking | 24% | default | — |
Pricing by serving provider
availabilityNo public price yet. Not offered on-demand at any tracked provider.
Price history
input + output $/1M since we started tracking Input Output
Price history is accruing. We record a point each day a price changes; the full backfill lands shortly.
Other Meta models
compare pricing across the Meta lineupLlama 3.3 70B Instruct$0.12/1M · 13 providersLlama 4 Scout 17B 16e Instruct$0.05/1M · 10 providersLlama 3.1 8B Instruct$0.02/1M · 8 providersMeta Llama 3.1 8B Instruct$0.02/1M · 6 providersLlama 3.2 3B Instruct$0.02/1M · 6 providersLlama 4 Maverick 17B 128e Instruct FP8$0.05/1M · 6 providersLlama 3.2 11B Vision Instruct$0.05/1M · 5 providersMeta Llama 3.1 70B Instruct$0.12/1M · 5 providers
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily
Every weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI