StackPricing
All AI models

AI model APIs · head-to-head

DeepSeek V4 Flash vs Gemini 2.5 Flash-Lite

Same usage, official rates, ranked by nothing but price. Both rate cards, the monthly bill at five workloads, and the exact usage mix where the answer flips.

The pricing verdict Gemini 2.5 Flash-Lite is cheaper than DeepSeek V4 Flash at every input:output mix — there is no usage pattern where DeepSeek V4 Flash costs less.

Rate cards

Per 1M tokens DeepSeek V4 Flash Gemini 2.5 Flash-Lite
Input $0.44 $0.10
Cached input $0.014 $0.01
Output $1.32 $0.40
Context window

Monthly cost at five workloads

Workload Tokens / month DeepSeek V4 Flash Gemini 2.5 Flash-Lite
Prototype 1M in · 0.3M out $0.84 $0.22
Small app 10M in · 3M out $8.36 $2.20
Production app 50M in · 15M out $41.8 $11
High volume 200M in · 60M out $167 $44
Heavy platform 1B in · 300M out $836 $220

Standard on-demand rates from each provider's official pricing page, verified July 2026; figures exclude caching and batch discounts and any long-context surcharges. DeepSeek V4 Flash: Cache-hit input bills at $0.014 per 1M at peak ($0.007 off-peak). The legacy deepseek-chat and deepseek-reasoner names map to this model. Figures use PEAK rates. DeepSeek bills by time of day: peak is 01:00-04:00 and 06:00-10:00 UTC, and off-peak (the other 17 hours) is half these prices. We publish peak so the figure can never understate a bill; halve it for off-peak traffic. Gemini 2.5 Flash-Lite: Text/image/video input rate; audio input bills at $0.30 per 1M. Price is one input — quality, latency and context limits differ. Confirm live pricing before committing.

Price both at your exact usage DeepSeek V4 Flash pricing Gemini 2.5 Flash-Lite pricing