Mistral · AI model API pricing
How much does Mistral Medium 3.5 cost?
Mistral's strongest general model — matches Gemini 3.5 Flash on input and beats it on output, from a European provider.
Rate card
| Input tokens | $1.5 / 1M |
| Cached input | $0.15 / 1M |
| Output tokens | $7.5 / 1M |
| Context window | — |
What that means per month
| Workload | Tokens / month | Cost / month |
|---|---|---|
| Prototype | 1M in · 0.3M out | $3.75/mo |
| Small app | 10M in · 3M out | $37.5/mo |
| Production app | 50M in · 15M out | $188/mo |
| High volume | 200M in · 60M out | $750/mo |
| Heavy platform | 1B in · 300M out | $3,750/mo |
Standard on-demand rates from Mistral's official pricing page, last checked 2026-07-27. Cached input reads bill at 10% of the input rate, with no cache-creation or storage fee. Monthly figures exclude caching and batch discounts. Confirm live pricing before committing.
Where Mistral Medium 3.5 sits on price
At the reference month of 10M input and 3M output tokens, Mistral Medium 3.5 costs $37.5, which makes it the 19th cheapest of the 30 models tracked here. That is 22× the price of Qwen-Flash, the cheapest model at this mix ($1.70). Among the 11 mid-tier models it ranks 8th, behind MiniMax M3 at $6.60.
Cache economics
Cached input on Mistral Medium 3.5 bills at $0.15 per 1M — 10% of its own input rate, against a 10% median across the models here that publish a cached-read price. 2 models discount cached reads more steeply than this. On the reference month, serving 80% of that input from cache takes the input side of the bill from $15 to $4.20, leaving output untouched at $22.5 — which is why a high reuse rate changes the ranking above rather than just shaving the total.
Cache reads only. Writes are billed separately by some providers and are excluded here — see the rate note above for Mistral's terms. Try it at your own hit rate on the AI cost calculator.
Cheaper alternatives
At a reference workload of 10M input / 3M output tokens a month, these cost less than Mistral Medium 3.5 ($37.5/mo):
-
Qwen3-Max $30/mo -
GLM-5.2 $27.2/mo -
DeepSeek V4 Pro $25.1/mo