Context
128K
Max output
8K
Serving providers
1
Cheapest input
$0.00 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.
Pricing by serving provider
standard tier · per 1M tokens · click a provider to expand| Serving provider | Input /1M | Output /1M | Endpoints | ||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| LemonadeCheapest | $0.00 | $0.00 | 1 | ||||||||||||||||
| |||||||||||||||||||
Price history
input + output $/1M since we started tracking Input Output
Cost calculator
estimate your monthly spend on this model$0
estimated / month
Model IDs
copy the exact identifier for your platformlemonade/Gemma-3-4b-it-GGUF
Frequently asked questions
Gemma 3 4B It Gguf pricing, context and availabilityHow much does Gemma 3 4B It Gguf cost?
Gemma 3 4B It Gguf costs $0.00 per 1M input tokens and $0.00 per 1M output tokens at its cheapest provider via Lemonade.
What is the context window of Gemma 3 4B It Gguf?
Gemma 3 4B It Gguf accepts up to 128K tokens of context and can return up to 8K output tokens.
What can Gemma 3 4B It Gguf do?
Gemma 3 4B It Gguf supports tool use and structured output.
Other Google models
compare pricing across the Google lineup Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily
Every weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI