High-volume product features

Gemini 2.5 Flash vs GPT-4o mini Cost

Side-by-side cost comparison of Gemini 2.5 Flash and GPT-4o mini for high-volume, low-latency apps.

Google

Gemini 2.5 Flash

Provider
Google
Modality
text
Context
1,000,000
Input / 1M
$0.30
Output / 1M
$2.50
Cached / 1M
$0.075
Batch discount
50%
View model page →

OpenAI

GPT-4o mini

Provider
OpenAI
Modality
text
Context
128,000
Input / 1M
$0.15
Output / 1M
$0.60
Cached / 1M
$0.075
Batch discount
50%
Training / 1M
$3.00
View model page →

Sample workload

Estimated spend on a shared scenario

Catalog rates updated 2026-07-31. Same inputs on both models for an apples-to-apples read.

Gemini 2.5 Flash

10k × 2k/500 tokens

$18.50

GPT-4o mini

10k × 2k/500 tokens

$6.00

Lower cost

GPT-4o mini is lower by $12.50 on this sample workload.

Keep comparing

More matchups

Approximate static list prices for planning only. Always verify against each provider’s official pricing page before production budgeting.