Context
262K
Max output
256K
Serving providers
3
Cheapest input
$0.20 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.
Benchmarks
independent evaluations · source: Artificial Analysis and Hugging Face leaderboardsSpeed & latency
Output speed
88 t/s
tokens / second
Time to first token
2.75s
latency
Headline indices
Individual benchmarks · click a row for effort variants
| Benchmark | Best score | Effort | Cost / task | |
|---|---|---|---|---|
| APEX Agents | 15% | default | — | |
| Automation Bench | 0.1 | default | — | |
| CritPt | 2% | default | — | |
| EvalComputeProxy | 435.3 | default | — | |
| GDPval | 1017.7 | default | — | |
| GPQA Diamond | 81% | default | — | |
| Harvey Lab | 0.7 | default | — | |
| Humanity's Last Exam | 21% | default | — | |
| IFBench | 67% | default | — | |
| IT-Bench SRE | 30% | default | — | |
| Long-Context Reasoning | 70% | default | — | |
| MMMU-Pro | 75% | default | — | |
| Omniscience | -37.3 | default | — | |
| Omniscience Accuracy | 0.3 | default | — | |
| Omniscience Non Hallucination | 0.2 | default | — | |
| SciCode | 40% | default | — | |
| SWE-bench Pro | 56% | default | — | |
| TerminalBench Hard | 36% | default | — | |
| TerminalBench v2.1 | 39% | default | — | |
| τ-bench Banking | 12% | default | — | |
| τ²-bench | 99% | default | — |
Pricing by serving provider
standard tier · per 1M tokens · click a provider to expand| Serving provider | Input /1M | Output /1M | Endpoints | |||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| NovitaCheapest | $0.20 | $1.15 | 1 | |||||||||||||
| ||||||||||||||||
| DeepInfraCheapest | $0.20 | $1.15 | 1 | |||||||||||||
| ||||||||||||||||
| OpenRouterCheapest | $0.20 | $1.15 | 1 | |||||||||||||
| ||||||||||||||||
Price history
input + output $/1M since we started tracking Input Output
Cost calculator
estimate your monthly spend on this model$132
estimated / month
Model IDs
copy the exact identifier for your platformnovita/stepfun/step-3.7-flashdeepinfra/stepfun-ai/Step-3.7-Flashstepfun/step-3.7-flash
Frequently asked questions
Step 3.7 Flash pricing, context and availabilityHow much does Step 3.7 Flash cost?
Step 3.7 Flash costs $0.20 per 1M input tokens and $1.15 per 1M output tokens at its cheapest provider via Novita.
What is the context window of Step 3.7 Flash?
Step 3.7 Flash accepts up to 262K tokens of context and can return up to 256K output tokens.
Which providers serve Step 3.7 Flash?
Step 3.7 Flash is available from 3 serving providers, each with its own pricing and model ID. Novita is currently the cheapest.
What can Step 3.7 Flash do?
Step 3.7 Flash supports image input (vision), tool use, reasoning, prompt caching and structured output.
Other Other models
compare pricing across the Other lineup Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily
Every weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI