Ministral 3 14B vs Gemma 4 31B
Small open-weight models that businesses consider for self-hosting. Both list prices from the makers’ own pages, and what three everyday jobs cost on each.
US dollars, standard rate, shortest context band. Price only: this page does not rank quality.
- Ministral 3 14B, per M tokens
- $0.20 / $0.20
- Gemma 4 31B, per M tokens
- No list price
3,000 tokens in and 400 out each, at list price.
Per million tokens, Ministral 3 14B costs $0.20 in and $0.20 out. Gemma 4 31B has no list price: Google publishes its weights but sells no API for it, so the cost depends on the host. Both read up to 256K tokens at once.
Side by side
Each maker’s published rate for its standard tier. Cached input is what a repeated prompt prefix costs once the provider has stored it.
| Fact | Ministral 3 14B | Gemma 4 31B |
|---|---|---|
| Maker | Mistral | |
| Tier | Small and fast | Small and fast |
| Input, per M tokens | $0.20 | Not published |
| Output, per M tokens | $0.20 | Not published |
| Cached input | Not published | Not published |
| Context window | 256K | 256K |
| Released | 2 Dec 2025 | 2 Apr 2026 |
| Weights | Open | Open |
The small print
What each base rate leaves out, from the makers’ own pages: surcharges, discounts, promotions and cache pricing.
- Ministral 3 14B
- API id ministral-14b-2512. Apache 2.0. Cached price not published. Batch 50% off. Mistral’s pricing page
- Gemma 4 31B
- No first-party paid price: the Gemini API pricing page lists Gemma 4 on the free tier only, with the paid tier 'Not available'. Open weights; cost depends on the third-party host. Context from https://ai.google.dev/gemma/docs/core/model_card_4 (256K tokens). Google’s pricing page
Other comparisons
Every head to head on the list that involves Ministral 3 14B or Gemma 4 31B.
Pay a quarter of list price
AIGROW API credit is metered at the makers’ list prices, and $25 buys $100 of usage. The credit page lists the models it covers; ask about any other before buying.
Prices checked 22 September 2026