GPT-5 mini vs Gemini 2.5 Flash

GPT-5 mini (OpenAI) costs $0.25 per million input tokens and $2.00 per million output tokens; Gemini 2.5 Flash (Google) costs $0.30 and $2.50. GPT-5 mini is 1.2× cheaper for a typical 3:1 input-to-output mix. Gemini 2.5 Flash has the larger context window (1.05M vs 272K tokens).

GPT-5 miniGemini 2.5 Flash
ProviderOpenAIGoogle
Input / 1M tokens$0.25$0.30
Output / 1M tokens$2.00$2.50
Cached input / 1M$0.025$0.03
Blended (3:1)$0.69$0.85
Context window272K1.05M
Max output128K66K
VisionYesYes
Tool useYesYes
ReasoningYesYes
Listed sinceAug 10, 2025Jun 20, 2025

Cost per 1,000 requests

WorkloadGPT-5 miniGemini 2.5 Flash
Chatbot reply2,000 input + 500 output tokens$1.50$1.85
RAG / search answer8,000 input + 700 output tokens$3.40$4.15
Long-document summary100,000 input + 2,000 output tokens$29.00$35.00

Blended price history

GPT-5 miniGemini 2.5 Flash
$0$0.250$0.500$0.750$1.00Jul 25Oct 25Jan 26Apr 26Jul 26Oct 26$0.688$0.850

Blended = (3 × input + output) ÷ 4, per million tokens. Each line starts when the model first appeared in the price list.

The companies

List prices for each provider's own API (USD), taken daily from LiteLLM's public model price list. Batch, caching and committed-use discounts aren't reflected. Price alone says nothing about output quality; check benchmarks and your own evaluations before switching.