AI model APIs · head-to-head
DeepSeek V4 Flash vs Gemini 2.5 Flash-Lite
Same usage, official rates, ranked by nothing but price. Both rate cards, the monthly bill at five workloads, and the exact usage mix where the answer flips.
The pricing verdict Gemini 2.5 Flash-Lite is cheaper than DeepSeek V4 Flash at every input:output mix — there is no usage pattern where DeepSeek V4 Flash costs less.
Rate cards
| Per 1M tokens | DeepSeek V4 Flash | Gemini 2.5 Flash-Lite |
|---|---|---|
| Input | $0.44 | $0.10 |
| Cached input | $0.014 | $0.01 |
| Output | $1.32 | $0.40 |
| Context window | — | — |
Monthly cost at five workloads
| Workload | Tokens / month | DeepSeek V4 Flash | Gemini 2.5 Flash-Lite |
|---|---|---|---|
| Prototype | 1M in · 0.3M out | $0.84 | $0.22 |
| Small app | 10M in · 3M out | $8.36 | $2.20 |
| Production app | 50M in · 15M out | $41.8 | $11 |
| High volume | 200M in · 60M out | $167 | $44 |
| Heavy platform | 1B in · 300M out | $836 | $220 |
Standard on-demand rates from each provider's official pricing page, verified July 2026; figures exclude caching and batch discounts and any long-context surcharges. DeepSeek V4 Flash: Cache-hit input bills at $0.014 per 1M at peak ($0.007 off-peak). The legacy deepseek-chat and deepseek-reasoner names map to this model. Figures use PEAK rates. DeepSeek bills by time of day: peak is 01:00-04:00 and 06:00-10:00 UTC, and off-peak (the other 17 hours) is half these prices. We publish peak so the figure can never understate a bill; halve it for off-peak traffic. Gemini 2.5 Flash-Lite: Text/image/video input rate; audio input bills at $0.30 per 1M. Price is one input — quality, latency and context limits differ. Confirm live pricing before committing.