Google's legacy Flash model, providing baseline speed and intelligence. Built by Google.
Prices updated
Input price
$0.50
per 1M tokens · Standard
Output price
$3.00
per 1M tokens · Standard
Input limit
1M
tokens
Output limit
66K
tokens
Input formats
Output formats
Gemini 3 Flash (Preview) price history
2 price records since 9/25/2026
Gemini 3 Flash (Preview) cost calculator
$25/month
Standard pricing
Overview
What is Gemini 3 Flash (Preview)?
Gemini 3 Flash (Preview) is a text & reasoning and vision and code & slms model from Google. Google's legacy Flash model, providing baseline speed and intelligence. Its 1M context window and $0.50 input price make it a candidate for quality-focused production applications.
Benchmarks
Gemini 3 Flash (Preview) benchmarks & speed
How Gemini 3 Flash (Preview) scores on standardized evaluations, and where it lands among every model we track.
17.9
Intelligence Index
195tok/s
Output speed
1193
DesignArena Elo
Reasoning
GPQA Diamond
81.2%
Graduate-level scientific reasoning · top 53%
HLE
15.0%
Humanity's Last Exam · top 57%
Latency & design
- Time to first token
- 0.86s
- DesignArena win rate
- 57.6%
- Design battles judged
- 4,414
Independent scores from Artificial Analysis and DesignArena · updated 10/8/2026. Higher is better; ranks compare against every model we track with that score.
Rates
Gemini 3 Flash (Preview) pricing
Pricing is based on token usage. Provider-specific caching, batch, regional, and tool charges may affect the final cost.
Input · Standard
$0.50
Per 1M tokens
Cached input · Standard
$0.05
Per 1M tokens
Capabilities
Gemini 3 Flash (Preview) Tools
Tools available when using Gemini 3 Flash (Preview) through supported provider APIs.
Function calling
Supported
Structured outputs
Supported
JSON mode
Supported
Reasoning
Supported
Built-in web search
Not supported
Log probabilities
Not supported
Deterministic seed
Supported
Parallel tool calls
Not supported
Prompt caching
Supported
Where to run it
Gemini 3 Flash (Preview) API providers
6 providers serve Gemini 3 Flash (Preview). Prices are per 1M tokens; uptime is the last 24 hours.
| Provider | Input | Output | Cached | Context | Uptime |
|---|---|---|---|---|---|
| $0.25 | $1.50 | $0.03 | 1M | 99.59% | |
| Google AI Studio | $0.25 | $1.50 | $0.03 | 1M | 99.99% |
| Google AI Studio | $0.50 | $3.00 | $0.05 | 1M | 99.95% |
| $0.50 | $3.00 | $0.05 | 1M | 99.63% | |
| $0.90 | $5.40 | $0.09 | 1M | 99.81% | |
| Google AI Studio | $0.90 | $5.40 | $0.09 | 1M | 99.94% |
- Input
- $0.25
- Output
- $1.50
- Cached
- $0.03
- Context
- 1M
- Uptime
- 99.59%
- Input
- $0.25
- Output
- $1.50
- Cached
- $0.03
- Context
- 1M
- Uptime
- 99.99%
- Input
- $0.50
- Output
- $3.00
- Cached
- $0.05
- Context
- 1M
- Uptime
- 99.95%
- Input
- $0.50
- Output
- $3.00
- Cached
- $0.05
- Context
- 1M
- Uptime
- 99.63%
- Input
- $0.90
- Output
- $5.40
- Cached
- $0.09
- Context
- 1M
- Uptime
- 99.81%
- Input
- $0.90
- Output
- $5.40
- Cached
- $0.09
- Context
- 1M
- Uptime
- 99.94%
Strengths and limitations
Strengths
- Measured results (AA Intelligence 17.9 · DesignArena Elo 1193 · Arena Elo 1222 · MMLU-Pro 77.3%) position it for demanding production work.
- 1M context supports large documents, repositories and extended conversations.
- Supports text & reasoning and vision and code & slms workloads in one model.
Limitations
- Generated tokens cost 6× more than input tokens, which matters for verbose responses.
- Large context capacity does not guarantee consistent retrieval across the entire prompt.
- Arena Elo and MMLU-Pro are directional; test accuracy, latency and reliability on your own workload before committing.
Arena Elo, MMLU-Pro and pricing figures are illustrative; validate current vendor terms before purchase.
Same provider
