Step 3.5 Flash
Context
n/a
Max output
n/a
Serving providers
0
Cheapest input
n/a /1M
Benchmarks
independent evaluations · source: Artificial AnalysisSpeed & latency
Output speed
235 t/s
tokens / second
Time to first token
1.17s
latency
Headline indices
Individual benchmarks · click a row for effort variants
| Benchmark | Best score | Effort | Cost / task | |
|---|---|---|---|---|
| CritPt | 2% | default | — | |
| EvalComputeProxy | 400.6 | default | — | |
| GPQA Diamond | 83% | default | — | |
| Humanity's Last Exam | 21% | default | — | |
| IFBench | 65% | default | — | |
| Long-Context Reasoning | 48% | default | — | |
| Omniscience | -41.8 | default | — | |
| Omniscience Accuracy | 0.2 | default | — | |
| Omniscience Non Hallucination | 0.1 | default | — | |
| SciCode | 40% | default | — | |
| TerminalBench Hard | 27% | default | — | |
| τ²-bench | 94% | default | — |
Pricing by serving provider
availabilityNo public price yet. Not offered on-demand at any tracked provider.
Price history
input + output $/1M since we started tracking Input Output
Price history is accruing. We record a point each day a price changes; the full backfill lands shortly.
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily
Every weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI