Gemini 2.5 Flash API pricing
As of 5 August 2026, Gemini 2.5 Flash costs $0.30 per 1M input tokens and $2.50 per 1M output tokens, and scores 55.1% on coding benchmarks (best of Aider polyglot and SWE-bench Verified).
Prices updated 5 August 2026
Gemini 2.5 Flash price by host
The same model weights are served by 2 providers. Prices are USD per 1M tokens; the cheapest output price is highlighted.
| Provider | Input / 1M | Output / 1M | Context |
|---|---|---|---|
| DeepInfra | $0.30 | $2.50 | 1M |
| Google (Gemini) | $0.30 | $2.50 | 1.0M |
Coding via Aider polyglot & SWE-bench Verified; reasoning & math via Epoch AI.
Cheaper alternatives to Gemini 2.5 Flash
Models with an equal-or-better coding score at a lower output price. StackSpend surfaces swaps like these against your actual usage.
- How much does the Gemini 2.5 Flash API cost?
- As of 5 August 2026, Gemini 2.5 Flash costs $0.30 per 1M input tokens and $2.50 per 1M output tokens, and scores 55.1% on coding benchmarks (best of Aider polyglot and SWE-bench Verified).
- What's the cheapest way to run Gemini 2.5 Flash?
- Gemini 2.5 Flash is served by 2 hosts. The lowest output price is $2.50 per 1M tokens via DeepInfra — the same weights, so switching host is a like-for-like saving.
- Is there a cheaper model as good as Gemini 2.5 Flash for coding?
- Yes — deepseek-v4-pro (DeepSeek) scores 77.6% on coding at $0.87 per 1M output tokens, below Gemini 2.5 Flash's $2.50.
Know what you're really spending on Gemini 2.5 Flash.
StackSpend connects read-only to your AI providers, normalises every model into one daily signal, and flags cheaper equal-quality swaps from this same dataset.