PriceIndex
Google logo

Gemma 4 31B

gemma-4-31b-it

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function... Built by Google.

Prices updated

Input price

$0.09

per 1M tokens · Standard

Output price

$0.34

per 1M tokens · Standard

Input limit

262K

tokens

Output limit

16K

tokens

Input formats

TextImageVideoAudioPDF

Output formats

TextImageVideoAudioPDF
Compare

Gemma 4 31B price history

2 price records since 9/25/2026

Gemma 4 31B cost calculator

Input20M tokens
Output5M tokens

$3.50/month

Standard pricing

    Overview

    What is Gemma 4 31B?

    Gemma 4 31B is a text & reasoning and vision model from Google. Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function... Its 262K context window and $0.09 input price make it a candidate for cost-sensitive, high-throughput applications.

    Image understandingDocument analysisContent workflowsClassification

    Benchmarks

    Gemma 4 31B benchmarks & speed

    How Gemma 4 31B scores on standardized evaluations, and where it lands among every model we track.

    14.7

    Intelligence Index

    Better than 37% of 153 models

    36tok/s

    Output speed

    Better than 1% of 119 models

    Reasoning

    GPQA Diamond

    85.7%

    Graduate-level scientific reasoning · top 41%

    HLE

    23.6%

    Humanity's Last Exam · top 46%

    Coding

    SciCode

    45.5%

    Python for scientific computing · top 68%

    Latency & design

    Time to first token
    1.12s

    Independent scores from Artificial Analysis and DesignArena · updated 10/5/2026. Higher is better; ranks compare against every model we track with that score.

    Rates

    Gemma 4 31B pricing

    Pricing is based on token usage. Provider-specific caching, batch, regional, and tool charges may affect the final cost.

    Input · Standard

    $0.09

    Per 1M tokens

    Cached input · Standard

    $0.05

    Per 1M tokens

    Capabilities

    Gemma 4 31B Tools

    Tools available when using Gemma 4 31B through supported provider APIs.

    Function calling

    Supported

    Structured outputs

    Supported

    JSON mode

    Supported

    Reasoning

    Supported

    Built-in web search

    Not supported

    Log probabilities

    Supported

    Deterministic seed

    Supported

    Parallel tool calls

    Not supported

    Prompt caching

    Supported

    Strengths and limitations

    Strengths

    • Low input cost at $0.09 per million tokens suits high-volume workloads.
    • 262K context supports large documents, repositories and extended conversations.
    • Supports text & reasoning and vision workloads in one model.

    Limitations

    • Generated tokens cost 4× more than input tokens, which matters for verbose responses.
    • Large context capacity does not guarantee consistent retrieval across the entire prompt.
    • Arena Elo and MMLU-Pro are directional; test accuracy, latency and reliability on your own workload before committing.

    Arena Elo, MMLU-Pro and pricing figures are illustrative; validate current vendor terms before purchase.

    Same provider

    More models from Google

    View provider →

    Frequently asked questions about Gemma 4 31B

    Gemma 4 31B is a text & reasoning and vision model from Google. Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...