DeepSeek V4 Flash vs Gemini 3.5 Flash-Lite

DeepSeek V4 Flash (DeepSeek) costs $0.30 per million input tokens and $1.20 per million output tokens; Gemini 3.5 Flash-Lite (Google) costs $0.30 and $2.50. DeepSeek V4 Flash is 1.6× cheaper for a typical 3:1 input-to-output mix. Gemini 3.5 Flash-Lite has the larger context window (1.05M vs 1M tokens).

DeepSeek V4 FlashGemini 3.5 Flash-Lite
ProviderDeepSeekGoogle
Input / 1M tokens$0.30$0.30
Output / 1M tokens$1.20$2.50
Cached input / 1M$0.006$0.03
Blended (3:1)$0.52$0.85
Context window1M1.05M
Max output393K66K
VisionYesYes
Tool useYesYes
ReasoningYesYes
Listed sinceJun 20, 2026Jul 25, 2026

Cost per 1,000 requests

WorkloadDeepSeek V4 FlashGemini 3.5 Flash-Lite
Chatbot reply2,000 input + 500 output tokens$1.20$1.85
RAG / search answer8,000 input + 700 output tokens$3.24$4.15
Long-document summary100,000 input + 2,000 output tokens$32.40$35.00

Blended price history

DeepSeek V4 FlashGemini 3.5 Flash-Lite
$0$0.250$0.500$0.750$1.00Jul 26Aug 26Sep 26Oct 26$0.525$0.850

Blended = (3 × input + output) ÷ 4, per million tokens. Each line starts when the model first appeared in the price list.

The companies

Related comparisons

List prices for each provider's own API (USD), taken daily from LiteLLM's public model price list. Batch, caching and committed-use discounts aren't reflected. Price alone says nothing about output quality; check benchmarks and your own evaluations before switching.