DeepSeek V4.1 Flash vs Gemini 3.5 Flash-Lite
Budget models with 1M context, compared for document-heavy tasks. Both list prices from the makers’ own pages, and what three everyday jobs cost on each.
US dollars, standard rate, shortest context band. Price only: this page does not rank quality.
- DeepSeek V4.1 Flash
- $1.38
- Gemini 3.5 Flash-Lite
- $1.90
- DeepSeek V4.1 Flash, per M tokens
- $0.30 / $1.20
- Gemini 3.5 Flash-Lite, per M tokens
- $0.30 / $2.50
3,000 tokens in and 400 out each, at list price.
Per million tokens, DeepSeek V4.1 Flash costs $0.30 in and $1.20 out; Gemini 3.5 Flash-Lite costs $0.30 in and $2.50 out. 1,000 support replies cost $1.90 on Gemini 3.5 Flash-Lite and $1.38 on DeepSeek V4.1 Flash, so Gemini 3.5 Flash-Lite costs 38% more. DeepSeek V4.1 Flash reads up to 1M tokens at once, Gemini 3.5 Flash-Lite up to 1.05M.
Side by side
Each maker’s published rate for its standard tier. Cached input is what a repeated prompt prefix costs once the provider has stored it.
| Fact | DeepSeek V4.1 Flash | Gemini 3.5 Flash-Lite |
|---|---|---|
| Maker | DeepSeek | |
| Tier | Small and fast | Small and fast |
| Input, per M tokens | $0.30 | $0.30 |
| Output, per M tokens | $1.20 | $2.50 |
| Cached input | $0.006 | $0.03 |
| Context window | 1M | 1.05M |
| Released | 10 Sep 2026 | 21 Jul 2026 |
| Weights | Open | Closed |
What a job costs on each
1,000 runs of each job at the two list prices. The gap moves from job to job because the models price input, output and cached context differently.
| Job | DeepSeek V4.1 Flash | Gemini 3.5 Flash-Lite | Cheaper |
|---|---|---|---|
| Support reply3,000 in, 400 out. | $1.38 | $1.90 | DeepSeek V4.1 Flash, 27% less |
| Document summary25,000 in, 1,000 out. | $8.70 | $10.00 | DeepSeek V4.1 Flash, 13% less |
| Agent task80,000 in (70,000 cached), 3,000 out. | $7.02 | $12.60 | DeepSeek V4.1 Flash, 44% less |
The token counts are ours, chosen to look like real work; the prices are the makers’. A model with no published cache price bills the agent task’s repeated context at its full input rate.
The small print
What each base rate leaves out, from the makers’ own pages: surcharges, discounts, promotions and cache pricing.
- DeepSeek V4.1 Flash
- Peak-hour rate, recorded as the list price. Off-peak is half: $0.15 / $0.60, cache hit $0.003. API name deepseek-flash. Max output 384K. DeepSeek’s pricing page
- Gemini 3.5 Flash-Lite
- Same price for text, image, video and audio input. Batch 50% off. Cache storage $1.00 per MTok per hour. Google’s pricing page
Other comparisons
Every head to head on the list that involves DeepSeek V4.1 Flash or Gemini 3.5 Flash-Lite.
Run both for a quarter of list price
AIGROW API credit covers DeepSeek V4.1 Flash and Gemini 3.5 Flash-Lite and is metered at the list prices above: $25 buys $100 of usage.
Prices checked 22 September 2026