Gemini 2.5 Flash API Pricing 2026: $0.15/M Tokens
Complete per-million-token pricing for Gemini 2.5 Flash from Google, including cached and batch rates, cost-per-1K/1M/10M breakdowns, a monthly cost calculator, and a head-to-head comparison with GPT-4o-mini, Claude Haiku 3.5, DeepSeek-V3. For full cross-provider rankings, open the main calculator.
Gemini 2.5 Flash headline pricing
Input
$0.15 /M tokens
Output
$0.60 /M tokens
Cached input
$0.0375 /M (75% off)
Context · capabilities
1M context · Chat, Code, Vision · benchmark 7.5/10
Gemini 2.5 Flash cost per 1K, 1M and 10M tokens
Standard on-demand USD rates. Input assumes prompt tokens; output assumes generated tokens.
| Volume | Input cost | Output cost | Total (1:1) |
|---|---|---|---|
| 1K tokens | $0.0002 | $0.0006 | $0.0007 |
| 1M tokens | $0.15 | $0.60 | $0.75 |
| 10M tokens | $1.50 | $6.00 | $7.50 |
Gemini 2.5 Flash monthly cost calculator
$0.45
per month for 1M in + 500K out
$5.40
per year
Need to weigh it against other models? Use the full cross-provider calculator.
Gemini 2.5 Flash vs GPT-4o-mini vs Claude Haiku 3.5 vs DeepSeek-V3
Gemini 2.5 Flash ranks #2 of 4 by blended per-million-token price in this comparison. Lower input + output is cheaper.
| Model | Input $/M | Output $/M | In+Out | Context |
|---|---|---|---|---|
| GPT-4o-miniOpenAI | $0.15 | $0.60 | $0.75 | 128K |
| YOUGemini 2.5 FlashGoogle | $0.15 | $0.60 | $0.75 | 1M |
| DeepSeek-V3DeepSeek | $0.27 | $1.10 | $1.37 | 128K |
| Claude Haiku 3.5Anthropic | $0.80 | $4.00 | $4.80 | 200K |
Related pricing pages
Gemini 2.5 Flash pricing FAQ
How much does Gemini 2.5 Flash cost?
Gemini 2.5 Flash is priced at $0.15 per million input tokens and $0.60 per million output tokens on Google. At a typical 2:1 input-to-output ratio, that works out to roughly $0.30 per million blended tokens. Context window is 1M.
Is Gemini 2.5 Flash cheaper than GPT-4o-mini?
Gemini 2.5 Flash and GPT-4o-mini are within ~2% of each other on blended per-million-token price ($0.75 vs $0.75), so the choice should hinge on quality, context window, and feature fit rather than price.
Does Gemini 2.5 Flash support prompt caching?
Yes. Google serves repeated input prefixes on Gemini 2.5 Flash at $0.0375/M — a 75% discount versus the standard $0.15/M input rate. Caching is automatic for qualifying prefixes and is most valuable for long system prompts and recurring context.
Can I use Gemini 2.5 Flash with the Batch API?
No published batch discount is available for Gemini 2.5 Flash. For non-latency-sensitive workloads, compare against GPT-4o-mini or another model that offers a batch tier.