ColinBuilds.com · Official Google Gemini API pricing
Gemini 2.5 Flash
- Input / 1M
- $0.30
- Output / 1M
- $2.50
- Official Google context window
- 1,000,000 tokens
Google Gemini API gemini-2.5-flash standard paid tier; input $0.30/1M text, image, or video tokens, $1.00/1M audio input tokens, and output $2.50/1M tokens including thinking tokens. · as of July 4, 2026
Compare models
Same per-1M input/output rates as above. Enter your monthly token volumes to compare bills.
Gemini 2.5 Flash
- Input: 1M × $0.30 = $0.30
- Output: 1M × $2.50 = $2.50
- Total
- $2.80
GPT-4o
- Input: 1M × $2.50 = $2.50
- Output: 1M × $10.00 = $10.00
- Total
- $12.50
Difference
$9.70
GPT-4o costs $9.70 more than Gemini 2.5 Flash.
Popular comparisons
30 total featuring Gemini 2.5 Flash.
Gemini 2.5 Flash is Google’s price-performance Gemini model for low-latency, high-volume API work where the app still needs reasoning support and a long context window.
What this model is
Gemini 2.5 Flash is listed by Google as gemini-2.5-flash. Google describes it as a hybrid reasoning model with thinking budgets and a 1M token context window. It is positioned as a practical Gemini API option for builders who want a balance of speed, cost, multimodal input, and reasoning rather than the highest-end Gemini Pro route.
The model is useful for everyday assistants, document workflows, extraction, lightweight coding help, and high-volume product features where price matters. No public benchmark value is shown here unless it can be tied to an exact source for this model identity.
Pricing notes
Google’s official Gemini API pricing page lists Gemini 2.5 Flash standard paid-tier pricing at $0.30 per 1M input tokens for text, image, and video; $1.00 per 1M audio input tokens; and $2.50 per 1M output tokens, including thinking tokens.
The calculator on this page uses the standard non-batch text/image/video rates: $0.30 input and $2.50 output per 1M tokens. Google also lists separate prices for batch usage, context caching, grounding, audio input, and other tools, so those should be checked separately before estimating a full production bill.
Benchmarks and specs
Google’s Gemini API pricing page identifies the API model as gemini-2.5-flash and describes it as a hybrid reasoning model with a 1M token context window and thinking budgets.
The page does not use a parameter-count claim or a third-party leaderboard score here. Those numbers should only be shown when exact benchmark sources are matched to this exact model identity.
Best fit
Gemini 2.5 Flash is best for builders who need an official Google API model for practical product work: assistants, summarisation, extraction, multimodal prompts, and long-context tasks where Gemini Pro may be more than the job needs.
It is also a useful comparison point against cheaper small models because the official API price is low enough to matter in real app budgets while still coming from a major model maker.