Gemini 2.5 Flash API pricing

As of 5 August 2026, Gemini 2.5 Flash costs $0.30 per 1M input tokens and $2.50 per 1M output tokens, and scores 55.1% on coding benchmarks (best of Aider polyglot and SWE-bench Verified).

Prices updated 5 August 2026

Gemini 2.5 Flash price by host

The same model weights are served by 2 providers. Prices are USD per 1M tokens; the cheapest output price is highlighted.

ProviderInput / 1MOutput / 1MContext
DeepInfra$0.30$2.501M
Google (Gemini)$0.30$2.501.0M
Coding: 55.1%Math: 73.1%

Coding via Aider polyglot & SWE-bench Verified; reasoning & math via Epoch AI.

Cheaper alternatives to Gemini 2.5 Flash

Models with an equal-or-better coding score at a lower output price. StackSpend surfaces swaps like these against your actual usage.

How much does the Gemini 2.5 Flash API cost?
As of 5 August 2026, Gemini 2.5 Flash costs $0.30 per 1M input tokens and $2.50 per 1M output tokens, and scores 55.1% on coding benchmarks (best of Aider polyglot and SWE-bench Verified).
What's the cheapest way to run Gemini 2.5 Flash?
Gemini 2.5 Flash is served by 2 hosts. The lowest output price is $2.50 per 1M tokens via DeepInfra — the same weights, so switching host is a like-for-like saving.
Is there a cheaper model as good as Gemini 2.5 Flash for coding?
Yes — deepseek-v4-pro (DeepSeek) scores 77.6% on coding at $0.87 per 1M output tokens, below Gemini 2.5 Flash's $2.50.

Know what you're really spending on Gemini 2.5 Flash.

StackSpend connects read-only to your AI providers, normalises every model into one daily signal, and flags cheaper equal-quality swaps from this same dataset.