Context
131K
Max output
n/a
Serving providers
1
Cheapest input
$0.27 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
DeepInfraCheapest$0.27$0.761
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$0.27$0.76deepinfra/google/gemma-4-31B-it-Ultra

Price history

input + output $/1M since we started tracking
Input Output

Price history is accruing. We record a point each day a price changes; the full backfill lands shortly.

Cost calculator

estimate your monthly spend on this model
$115
estimated / month

Model IDs

copy the exact identifier for your platform
deepinfra/google/gemma-4-31B-it-Ultra

Frequently asked questions

Gemma 4 31B It Ultra pricing, context and availability

How much does Gemma 4 31B It Ultra cost?

Gemma 4 31B It Ultra costs $0.27 per 1M input tokens and $0.76 per 1M output tokens at its cheapest provider via DeepInfra.

What is the context window of Gemma 4 31B It Ultra?

Gemma 4 31B It Ultra accepts up to 131K tokens of context.

What can Gemma 4 31B It Ultra do?

Gemma 4 31B It Ultra supports image input (vision), tool use, reasoning and structured output.

Other Google models

compare pricing across the Google lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI