Mistral · AI model API pricing
How much does Mistral Large 3 cost?
Open-weight and aggressively priced — cheaper than most budget tiers despite being a large model.
Rate card
| Input tokens | $0.50 / 1M |
| Cached input | $0.05 / 1M |
| Output tokens | $1.5 / 1M |
| Context window | — |
What that means per month
| Workload | Tokens / month | Cost / month |
|---|---|---|
| Prototype | 1M in · 0.3M out | $0.95/mo |
| Small app | 10M in · 3M out | $9.50/mo |
| Production app | 50M in · 15M out | $47.5/mo |
| High volume | 200M in · 60M out | $190/mo |
| Heavy platform | 1B in · 300M out | $950/mo |
Standard on-demand rates from Mistral's official pricing page, last checked 2026-07-27. Despite the name, Large 3 is priced below Medium 3.5 — it is Mistral's open-weight line; Medium 3.5 is the premium API model. Cached input reads bill at 10% of the input rate, with no cache-creation or storage fee. Monthly figures exclude caching and batch discounts. Confirm live pricing before committing.
Where Mistral Large 3 sits on price
At the reference month of 10M input and 3M output tokens, Mistral Large 3 costs $9.50, which makes it the 10th cheapest of the 30 models tracked here. That is 5.6× the price of Qwen-Flash, the cheapest model at this mix ($1.70). Among the 11 mid-tier models it ranks 3rd, behind MiniMax M3 at $6.60.
Cache economics
Cached input on Mistral Large 3 bills at $0.05 per 1M — 10% of its own input rate, against a 10% median across the models here that publish a cached-read price. 2 models discount cached reads more steeply than this. On the reference month, serving 80% of that input from cache takes the input side of the bill from $5 to $1.40, leaving output untouched at $4.50 — which is why a high reuse rate changes the ranking above rather than just shaving the total.
Cache reads only. Writes are billed separately by some providers and are excluded here — see the rate note above for Mistral's terms. Try it at your own hit rate on the AI cost calculator.
Cheaper alternatives
At a reference workload of 10M input / 3M output tokens a month, these cost less than Mistral Large 3 ($9.50/mo):
-
DeepSeek V4 Flash $8.36/mo -
Qwen-Plus $7.60/mo -
Gemini 3.1 Flash-Lite $7/mo