Mistral Small 4 API Pricing
Mistral Small 4 is the cheap multimodal Mistral tier, and it just got less cheap: $0.15/M input and $0.60/M output, against $0.10 / $0.30 before. Input is up 1.5x and output has doubled, with no announcement on the vendor's page - the card simply shows different numbers. It remains an Apache 2.0 open-weight model, described by Mistral as a hybrid unifying instruct, reasoning and coding.
Run the numbers.
Calculator pre-loaded with current Mistral Small 4 rates from Mistral's pricing page. Tweak spend or output mix to compare this model with nearby alternatives.
Real-world presets.
Repo-wide bug fix
Reading 100-page contracts
Support agent ticket triage
Research planning turn
Paste text. See tokens. See cost.
Counts use a chars-per-token calibration measured on the vendor's own published tokenizer (mistralai/Mistral-Nemo-Instruct-2407, 2026-06-10). English prose is typically within a few percent; code and non-Latin scripts tokenize heavier. For billing-exact counts use the vendor's count-tokens API.
| Model | Input /M | Output /M | Effective blended | Context | Best for |
|---|---|---|---|---|---|
| Mistral Small 4 | $0.15 cache $0.15 | $0.60 | $0.186 current page | 256K | Low-cost Mistral workloads |
| Mistral Large 3 | $0.50 | $1.50 | $0.58 pricier | documented elsewhere — not on pricing page | Open-weight multimodal work |
| Mistral Medium 3.5 | $1.50 | $7.50 | $1.98 pricier | documented elsewhere — not on pricing page | Agentic and coding workloads |
| Devstral 2 | $0.40 | $2.00 | $0.528 pricier | documented elsewhere — not on pricing page | Coding and software agents |
| GPT-5.4 mini | $0.75 cache $0.075 | $4.50 | $0.541 pricier | 400K | Open-weight multimodal work |
| Claude Sonnet 4.6 | $3.00 cache $0.30 | $15.00 | $2.05 pricier | 1M | Open-weight multimodal work |
By comparison
Frequently asked.
Short answers for teams checking Mistral Small 4 pricing, status, and migration choices.
Q · 01 Is Mistral Small 4 still available? +
mistral-small-latest.Q · 02 How much does Mistral Small 4 cost? +
Q · 03 Is cached-input pricing included? +
Q · 04 What should teams compare it against? +
current same-provider models.