Devstral Small 2 API Pricing
Devstral Small 2 is retired. Mistral gives LABS-DEVSTRAL-SMALL-2512 a retirement date of March 31, 2026 and names Mistral Small 4 as the alternative; it appears in neither the live catalogue nor the API pricing page (checked August 23, 2026). This archive preserves the last published card: $0.1/M input and $0.3/M output.
Run the numbers.
Live calculator pre-loaded with Devstral Small 2 rates. Tweak spend, output mix, or cache assumptions to compare it with sibling models.
Real-world presets.
Invoice extraction
Support assistant
Knowledge base answer
Image QA review
Paste text. See tokens. See cost.
Counts use a chars-per-token calibration measured on the vendor's own published tokenizer (mistralai/Mistral-Nemo-Instruct-2407, 2026-06-10). English prose is typically within a few percent; code and non-Latin scripts tokenize heavier. For billing-exact counts use the vendor's count-tokens API.
| Model | Input /M | Output /M | Effective blended | Context | Best for |
|---|---|---|---|---|---|
| Devstral Small 2 Current | $0.10 | $0.30 | $0.116 agentic 92/8 | documented elsewhere | Open coding agents |
| Mistral Large 3 | $0.50 | $1.50 | $0.58 pricier | documented elsewhere — not on pricing page | Mistral production workloads |
| Mistral Small 4 | $0.15 cache $0.15 | $0.60 | $0.186 pricier | 256K | Low-cost multimodal chat |
| Magistral Medium | $2.00 | $5.00 | $2.24 pricier | documented elsewhere — not on pricing page | Transparent reasoning |
| Magistral Small | $0.50 | $1.50 | $0.58 pricier | documented elsewhere | Transparent reasoning |
| Ministral 3 14B | $0.20 | $0.20 | $0.20 pricier | 128K | Mistral production workloads |
| Ministral 3 3B | $0.10 | $0.10 | $0.10 cheaper | 128K | Mistral production workloads |
| DeepSeek V4 Flash | $0.44 cache $0.014 | $1.32 | $0.189 cheaper | 1M | Mistral production workloads |
Frequently asked.
Practical pricing questions, separated from calculator assumptions.
Q · 01 What is Devstral Small 2 priced at? +
$0.1/M input and $0.3/M output. The page stores USD per-million-token baseline pricing from mistral.ai.Q · 02 Does this page include higher context pricing tiers? +
Q · 03 Is prompt caching priced separately? +
$0.1/M.Q · 04 How is the effective price calculated? +
$0.12/M with only documented cache discounts included.