Gemma 4 31B: open weights, no list price
Google publishes the weights of Gemma 4 31B but sells no API for it, so there is no list price to work from. What is known about it is below, beside the models it is weighed against.
US dollars per million tokens, standard rate, shortest context band.
- Context window
- 256,000 tokens
- Released
- 2 Apr 2026
- Weights
- Open
From Google’s own pages, 22 September 2026.
Gemma 4 31B is an open-weight model from Google, which does not sell it on a paid API of its own, so it has no list price: what it costs depends on the host that runs it, or on your own hardware. It reads up to 256,000 tokens at once and was released on 2 Apr 2026.
Gemma 4 31B on real small-business tasks
What the AIGROW benchmark measured when Gemma 4 31B was given 20 office tasks, each marked pass or fail by fixed checks.
- Success rate
- 60% of 20 runs passed (95% interval 39 to 78%), 27th of 36 models.
- Cost per 1,000 passes
- $0.558, failed runs included.
- Answer time
- 11 s median, 36 s at the 90th percentile.
- Strongest and weakest
- Email triage (100% passed) and review replies (0%).
The small print
What the base rate above leaves out, copied from the maker’s pages: surcharges, discounts, promotions and cache pricing.
No first-party paid price: the Gemini API pricing page lists Gemma 4 on the free tier only, with the paid tier 'Not available'. Open weights; cost depends on the third-party host. Context from https://ai.google.dev/gemma/docs/core/model_card_4 (256K tokens).
Gemma 4 31B against the models it is weighed with
Each comparison puts both list prices side by side and costs the same three jobs on each.
- vs Ministral 3 14B
- Small open-weight models that businesses consider for self-hosting.
- vs Muse Glimmer 30B
- Two open-weight models around 30B parameters meant for consumer GPUs and workstations.
- vs Qwen3.8 27B
- Mid-size open-weight models that businesses consider for private deployment, only one of which has a first-party API price.
Other Google models
| Model | Input | Output | Context | 1,000 replies |
|---|---|---|---|---|
| Gemini 3.1 Flash-LiteGoogle | $0.25 | $1.50 | 1.05M | $1.35 |
| Gemini 3.5 Flash-LiteGoogle | $0.30 | $2.50 | 1.05M | $1.90 |
| Gemini 3.8 FlashGoogle | $0.75 | $3.75 | 1.05M | $3.75 |
| Gemini 3.1 Pro (preview)Google | $2.00 | $12.00 | 1.05M | $10.80 |
Where the price comes from
Read on 22 September 2026. Each price is the standard rate for the shortest context band, with no batch or priority discount; long-context surcharges, batch rates and promotions are in each model’s small print. No price was taken from a reseller.
Pay a quarter of list price
AIGROW API credit is metered at the makers’ list prices, and $25 buys $100 of usage. The credit page lists the models it covers; ask about any other before buying.
Prices checked 22 September 2026