Gemini 3.1 Flash-Lite: what it costs to run
The price Google lists for Gemini 3.1 Flash-Lite, what a support reply, a document summary and an agent task cost at it, and where that puts it among 37 models.
US dollars per million tokens, standard rate, shortest context band.
- 1,000 support replies
- $1.35
- Cached input
- $0.025
- Context window
- 1,048,576 tokens
- Released
- 7 May 2026
- Weights
- Closed
From Google’s own pages, 22 September 2026.
Gemini 3.1 Flash-Lite costs $0.25 per million input tokens and $1.50 per million output tokens on Google’s own API, and $0.025 for cached input. At those prices 1,000 support replies cost $1.35, the 7th cheapest of 33 priced models on our list.
What a job costs
Three jobs, 1,000 runs of each, at the list price above, and where each puts Gemini 3.1 Flash-Lite among the 33 models with a list price.
| Job | 1,000 runs | Among 33 |
|---|---|---|
| Support reply3,000 in, 400 out. | $1.35 | 7th cheapest |
| Document summary25,000 in, 1,000 out. | $7.75 | 7th cheapest |
| Agent task80,000 in (70,000 cached), 3,000 out. | $8.75 | 5th cheapest |
The token counts are ours, chosen to look like real work; the prices are Google’s. All three jobs stay under 200,000 tokens a request, below every long-context surcharge on the list.
Gemini 3.1 Flash-Lite on real small-business tasks
What the AIGROW benchmark measured when Gemini 3.1 Flash-Lite was given 20 office tasks, each marked pass or fail by fixed checks.
- Success rate
- 45% of 20 runs passed (95% interval 26 to 66%), 28th of 36 models.
- Cost per 1,000 passes
- $1.41, failed runs included.
- Answer time
- 1.7 s median, 3.2 s at the 90th percentile.
- Strongest and weakest
- Email triage (100% passed) and review replies (0%).
The small print
What the base rate above leaves out, copied from the maker’s pages: surcharges, discounts, promotions and cache pricing.
Audio input $0.50 (cached $0.05). Batch 50% off. Cache storage $1.00 per MTok per hour.
Gemini 3.1 Flash-Lite against the models it is weighed with
Each comparison puts both list prices side by side and costs the same three jobs on each.
- vs GPT-5.6 Luna
- OpenAI's nano tier against Google's cheapest Gemini, close on price.
- vs Gemini 3.5 Flash-Lite
- Google's two Flash-Lite generations, where the newer one costs more on output.
Other Google models
| Model | Input | Output | Context | 1,000 replies |
|---|---|---|---|---|
| Gemini 3.5 Flash-LiteGoogle | $0.30 | $2.50 | 1.05M | $1.90 |
| Gemini 3.8 FlashGoogle | $0.75 | $3.75 | 1.05M | $3.75 |
| Gemini 3.1 Pro (preview)Google | $2.00 | $12.00 | 1.05M | $10.80 |
| Gemma 4 31BGoogle · open weights | No list price | · | 256K | · |
Where the price comes from
Read on 22 September 2026. Each price is the standard rate for the shortest context band, with no batch or priority discount; long-context surcharges, batch rates and promotions are in each model’s small print. No price was taken from a reseller.
Pay a quarter of list price
AIGROW API credit is metered at the makers’ list prices, and $25 buys $100 of usage. The credit page lists the models it covers; ask about any other before buying.
Prices checked 22 September 2026