Mistral Small 4: what it costs to run
The price Mistral lists for Mistral Small 4, what a support reply, a document summary and an agent task cost at it, and where that puts it among 37 models.
US dollars per million tokens, standard rate, shortest context band.
- 1,000 support replies
- $0.69
- Cached input
- Not published
- Context window
- 256,000 tokens
- Released
- 16 Mar 2026
- Weights
- Open
From Mistral’s own pages, 22 September 2026.
Mistral Small 4 costs $0.15 per million input tokens and $0.60 per million output tokens on Mistral’s own API. At those prices 1,000 support replies cost $0.69, the 4th cheapest of 33 priced models on our list. It reads up to 256,000 tokens at once and was released on 16 Mar 2026.
What a job costs
Three jobs, 1,000 runs of each, at the list price above, and where each puts Mistral Small 4 among the 33 models with a list price.
| Job | 1,000 runs | Among 33 |
|---|---|---|
| Support reply3,000 in, 400 out. | $0.69 | 4th cheapest |
| Document summary25,000 in, 1,000 out. | $4.35 | 3rd cheapest |
| Agent task80,000 in (70,000 cached), 3,000 out. | $13.80 | 9th cheapest |
The token counts are ours, chosen to look like real work; the prices are Mistral’s. All three jobs stay under 200,000 tokens a request, below every long-context surcharge on the list. Mistral publishes no cache price for this model, so the agent task bills its repeated context at the full input rate.
Mistral Small 4 on real small-business tasks
What the AIGROW benchmark measured when Mistral Small 4 was given 20 office tasks, each marked pass or fail by fixed checks.
- Success rate
- 20% of 20 runs passed (95% interval 8 to 42%), 35th of 36 models.
- Cost per 1,000 passes
- $1.33, failed runs included.
- Answer time
- 2.0 s median, 20 s at the 90th percentile.
- Strongest and weakest
- Request routing (100% passed) and review replies (0%).
The small print
What the base rate above leaves out, copied from the maker’s pages: surcharges, discounts, promotions and cache pricing.
API id mistral-small-2603. 119B total, 6.5B active parameters, Apache 2.0. Cached price not published. Batch 50% off.
Mistral Small 4 against the models it is weighed with
Each comparison puts both list prices side by side and costs the same three jobs on each.
- vs GPT-5.6 Luna
- Mistral's open-weight small model against OpenAI's nano tier, both near the bottom of the price range.
Other Mistral models
| Model | Input | Output | Context | 1,000 replies |
|---|---|---|---|---|
| Ministral 3 14BMistral · open weights | $0.20 | $0.20 | 256K | $0.68 |
| Mistral Medium 3.5Mistral · open weights | $1.50 | $7.50 | 256K | $7.50 |
Where the price comes from
Read on 22 September 2026. Each price is the standard rate for the shortest context band, with no batch or priority discount; long-context surcharges, batch rates and promotions are in each model’s small print. No price was taken from a reseller.
Pay a quarter of list price
AIGROW API credit is metered at the makers’ list prices, and $25 buys $100 of usage. The credit page lists the models it covers; ask about any other before buying.
Prices checked 22 September 2026