PriceIndex
Google logo

Gemini 2.5 Flash

gemini-2.5-flash

Fast, low-cost multimodal workhorse. Built by Google.

Prices updated

Input price

$0.30

per 1M tokens · Standard

Output price

$2.50

per 1M tokens · Standard

Input limit

1M

tokens

Output limit

66K

tokens

Input formats

TextImageVideoAudioPDF

Output formats

TextImageVideoAudioPDF
Compare

Gemini 2.5 Flash price history

12 price records since 4/1/2026 · 10 changes

Gemini 2.5 Flash cost calculator

Input20M tokens
Output5M tokens

$19/month

Standard pricing

Overview

What is Gemini 2.5 Flash?

Gemini 2.5 Flash is a text & reasoning and vision model from Google. Fast, low-cost multimodal workhorse. Its 1M context window and $0.30 input price make it a candidate for cost-sensitive, high-throughput applications.

Image understandingDocument analysisContent workflowsClassification

Benchmarks

Gemini 2.5 Flash benchmarks & speed

How Gemini 2.5 Flash scores on standardized evaluations, and where it lands among every model we track.

9.9

Intelligence Index

Better than 26% of 153 models

190tok/s

Output speed

Better than 85% of 119 models

1063

DesignArena Elo

Better than 15% of 81 models

Reasoning

GPQA Diamond

68.3%

Graduate-level scientific reasoning · top 72%

HLE

4.7%

Humanity's Last Exam · top 83%

Latency & design

Time to first token
0.46s
DesignArena win rate
45.4%
Design battles judged
6,971

Independent scores from Artificial Analysis and DesignArena · updated 10/8/2026. Higher is better; ranks compare against every model we track with that score.

Rates

Gemini 2.5 Flash pricing

Pricing is based on token usage. Provider-specific caching, batch, regional, and tool charges may affect the final cost.

Input · Standard

$0.30

Per 1M tokens

Cached input · Standard

$0.03

Per 1M tokens

Capabilities

Gemini 2.5 Flash Tools

Tools available when using Gemini 2.5 Flash through supported provider APIs.

Function calling

Supported

Structured outputs

Supported

JSON mode

Supported

Reasoning

Supported

Built-in web search

Not supported

Log probabilities

Not supported

Deterministic seed

Supported

Parallel tool calls

Not supported

Prompt caching

Supported

Where to run it

Gemini 2.5 Flash API providers

7 providers serve Gemini 2.5 Flash. Prices are per 1M tokens; uptime is the last 24 hours.

Google
Input
$0.30
Output
$2.50
Cached
$0.03
Context
1M
Uptime
95.58%
Google
Input
$0.30
Output
$2.50
Cached
$0.03
Context
1M
Uptime
99.16%
Google
Input
$0.54
Output
$4.50
Cached
$0.05
Context
1M
Uptime
99.9%
Google AI Studio
Input
$0.15
Output
$1.25
Cached
$0.01
Context
1M
Uptime
99.95%
Google AI Studio
Input
$0.30
Output
$2.50
Cached
$0.03
Context
1M
Uptime
99.84%
Google
Input
$0.30
Output
$2.50
Cached
$0.03
Context
1M
Uptime
90.7%

6 of 7

Strengths and limitations

Strengths

  • Low input cost at $0.30 per million tokens suits high-volume workloads.
  • 1M context supports large documents, repositories and extended conversations.
  • Supports text & reasoning and vision workloads in one model.

Limitations

  • Generated tokens cost 8× more than input tokens, which matters for verbose responses.
  • Large context capacity does not guarantee consistent retrieval across the entire prompt.
  • Arena Elo and MMLU-Pro are directional; test accuracy, latency and reliability on your own workload before committing.

Arena Elo, MMLU-Pro and pricing figures are illustrative; validate current vendor terms before purchase.

Same provider

More models from Google

View provider →

Peer set

Alternatives to Gemini 2.5 Flash

128K

$0.15 in · $0.60 out

Compare side-by-side
DeepSeek V3

DeepSeek

164K

$0.26 in · $1.03 out

Compare side-by-side

Frequently asked questions about Gemini 2.5 Flash

Gemini 2.5 Flash is a text & reasoning and vision model from Google. Fast, low-cost multimodal workhorse.