Context
131K
Max output
n/a
Serving providers
1
Cheapest input
$0.07 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: Artificial Analysis and Hugging Face leaderboards
Speed & latency
Output speed
319 t/s
tokens / second
Time to first token
7.25s
latency
Headline indices
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
AIME 202693%default
CritPt2%default
EvalComputeProxy1941.0default
GDPval1105.4default
GPQA Diamond85%default
Humanity's Last Exam22%default
Long-Context Reasoning60%default
Omniscience-16.9default
SciCode41%default
SWE-bench Pro57%default
TerminalBench v2.155%default
τ-bench Banking28%default

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
OpenRouterCheapest$0.07$0.221
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$0.07$0.22$0.01inclusionai/ling-3.0-flash

Price history

input + output $/1M since we started tracking
Input Output

Price history is accruing. We record a point each day a price changes; the full backfill lands shortly.

Cost calculator

estimate your monthly spend on this model
$33
estimated / month

Model IDs

copy the exact identifier for your platform
inclusionai/ling-3.0-flash

Frequently asked questions

Ling-3.0-flash pricing, context and availability

How much does Ling-3.0-flash cost?

Ling-3.0-flash costs $0.07 per 1M input tokens and $0.22 per 1M output tokens at its cheapest provider via OpenRouter.

What is the context window of Ling-3.0-flash?

Ling-3.0-flash accepts up to 131K tokens of context.

Other Other models

compare pricing across the Other lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI