Google · Updated 2026-07-22

Gemini 2.5 Flash API Pricing 2026: $0.15/M Tokens

Complete per-million-token pricing for Gemini 2.5 Flash from Google, including cached and batch rates, cost-per-1K/1M/10M breakdowns, a monthly cost calculator, and a head-to-head comparison with GPT-4o-mini, Claude Haiku 3.5, DeepSeek-V3. For full cross-provider rankings, open the main calculator.

Gemini 2.5 Flash headline pricing

Input

$0.15 /M tokens

Output

$0.60 /M tokens

Cached input

$0.0375 /M (75% off)

Context · capabilities

1M context · Chat, Code, Vision · benchmark 7.5/10

Gemini 2.5 Flash cost per 1K, 1M and 10M tokens

Standard on-demand USD rates. Input assumes prompt tokens; output assumes generated tokens.

VolumeInput costOutput costTotal (1:1)
1K tokens$0.0002$0.0006$0.0007
1M tokens$0.15$0.60$0.75
10M tokens$1.50$6.00$7.50

Gemini 2.5 Flash monthly cost calculator

1M
500K

$0.45

per month for 1M in + 500K out

$5.40

per year

Need to weigh it against other models? Use the full cross-provider calculator.

Gemini 2.5 Flash vs GPT-4o-mini vs Claude Haiku 3.5 vs DeepSeek-V3

Gemini 2.5 Flash ranks #2 of 4 by blended per-million-token price in this comparison. Lower input + output is cheaper.

ModelInput $/MOutput $/MIn+OutContext
GPT-4o-miniOpenAI$0.15$0.60$0.75128K
YOUGemini 2.5 FlashGoogle$0.15$0.60$0.751M
DeepSeek-V3DeepSeek$0.27$1.10$1.37128K
Claude Haiku 3.5Anthropic$0.80$4.00$4.80200K

Related pricing pages

Gemini 2.5 Flash pricing FAQ

How much does Gemini 2.5 Flash cost?

Gemini 2.5 Flash is priced at $0.15 per million input tokens and $0.60 per million output tokens on Google. At a typical 2:1 input-to-output ratio, that works out to roughly $0.30 per million blended tokens. Context window is 1M.

Is Gemini 2.5 Flash cheaper than GPT-4o-mini?

Gemini 2.5 Flash and GPT-4o-mini are within ~2% of each other on blended per-million-token price ($0.75 vs $0.75), so the choice should hinge on quality, context window, and feature fit rather than price.

Does Gemini 2.5 Flash support prompt caching?

Yes. Google serves repeated input prefixes on Gemini 2.5 Flash at $0.0375/M — a 75% discount versus the standard $0.15/M input rate. Caching is automatic for qualifying prefixes and is most valuable for long system prompts and recurring context.

Can I use Gemini 2.5 Flash with the Batch API?

No published batch discount is available for Gemini 2.5 Flash. For non-latency-sensitive workloads, compare against GPT-4o-mini or another model that offers a batch tier.

☕ Support this tool