Claude Haiku 4.5 vs Gemini 3.5 Flash-Lite

Claude Haiku 4.5 (Anthropic) costs $1.00 per million input tokens and $5.00 per million output tokens; Gemini 3.5 Flash-Lite (Google) costs $0.30 and $2.50. Gemini 3.5 Flash-Lite is 2.4× cheaper for a typical 3:1 input-to-output mix. Gemini 3.5 Flash-Lite has the larger context window (1.05M vs 200K tokens).

Claude Haiku 4.5Gemini 3.5 Flash-Lite
ProviderAnthropicGoogle
Input / 1M tokens$1.00$0.30
Output / 1M tokens$5.00$2.50
Cached input / 1M$0.10$0.03
Blended (3:1)$2.00$0.85
Context window200K1.05M
Max output64K66K
VisionYesYes
Tool useYesYes
ReasoningYesYes
Listed sinceOct 18, 2025Jul 25, 2026

Cost per 1,000 requests

WorkloadClaude Haiku 4.5Gemini 3.5 Flash-Lite
Chatbot reply2,000 input + 500 output tokens$4.50$1.85
RAG / search answer8,000 input + 700 output tokens$11.50$4.15
Long-document summary100,000 input + 2,000 output tokens$110.00$35.00

Blended price history

Claude Haiku 4.5Gemini 3.5 Flash-Lite
$0$0.600$1.20$1.80$2.40Jan 26Apr 26Jul 26Oct 26$2.00$0.850

Blended = (3 × input + output) ÷ 4, per million tokens. Each line starts when the model first appeared in the price list.

The companies

Related comparisons

List prices for each provider's own API (USD), taken daily from LiteLLM's public model price list. Batch, caching and committed-use discounts aren't reflected. Price alone says nothing about output quality; check benchmarks and your own evaluations before switching.