Context
262K
Max output
256K
Serving providers
5
Cheapest input
$0.07 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand

Input pricing runs from $0.07 to $0.25 per 1M tokens across 5 providers, so the dearest route costs 257% more than the cheapest for the same model.

Serving providerInput /1MOutput /1MEndpoints
DeepInfraCheapest$0.07$0.341
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$0.07$0.34deepinfra/google/gemma-4-26B-A4B-it
OpenRouterCheapest$0.07$0.341
Cloudflare$0.10$0.301
Novita$0.13$0.401
Scaleway$0.25$0.501

Price history

input + output $/1M since we started tracking
Input Output
Input down 72% since first tracked
$0.000$0.200$0.400$0.600Jun 12Jul 19Jul 29Aug 28$0.300$0.070

Cost calculator

estimate your monthly spend on this model
$41
estimated / month

Model IDs

copy the exact identifier for your platform
deepinfra/google/gemma-4-26B-A4B-itgoogle/gemma-4-26b-a4b-itcloudflare/@cf/google/gemma-4-26b-a4b-itnovita/google/gemma-4-26b-a4b-itscaleway/google/gemma-4-26b-a4b-it

Frequently asked questions

Gemma 4 26B A4B pricing, context and availability

How much does Gemma 4 26B A4B cost?

Gemma 4 26B A4B costs $0.07 per 1M input tokens and $0.30 per 1M output tokens at its cheapest provider via DeepInfra. Across 5 serving providers, input prices range from $0.07 to $0.25 per 1M tokens.

What is the context window of Gemma 4 26B A4B ?

Gemma 4 26B A4B accepts up to 262K tokens of context and can return up to 256K output tokens.

Which providers serve Gemma 4 26B A4B ?

Gemma 4 26B A4B is available from 5 serving providers, each with its own pricing and model ID. DeepInfra is currently the cheapest.

What can Gemma 4 26B A4B do?

Gemma 4 26B A4B supports image input (vision), tool use, reasoning and structured output.

Other Google models

compare pricing across the Google lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI