PriceIndex
IN

Mercury 2

mercury-2

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving... Built by Inception.

Prices updated

Input price

$0.25

per 1M tokens · Standard

Output price

$0.75

per 1M tokens · Standard

Input limit

128K

tokens

Output limit

50K

tokens

Input formats

TextImageVideoAudioPDF

Output formats

TextImageVideoAudioPDF
Compare

Mercury 2 price history

2 price records since 10/8/2026

Mercury 2 cost calculator

Input20M tokens
Output5M tokens

$8.75/month

Standard pricing

    Overview

    What is Mercury 2?

    Mercury 2 is a text & reasoning model from Inception. Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving... Its 128K context window and $0.25 input price make it a candidate for cost-sensitive, high-throughput applications.

    Content workflowsClassification

    Rates

    Mercury 2 pricing

    Pricing is based on token usage. Provider-specific caching, batch, regional, and tool charges may affect the final cost.

    Input · Standard

    $0.25

    Per 1M tokens

    Cached input · Standard

    $0.03

    Per 1M tokens

    Capabilities

    Mercury 2 Tools

    Tools available when using Mercury 2 through supported provider APIs.

    Function calling

    Supported

    Structured outputs

    Supported

    JSON mode

    Supported

    Reasoning

    Supported

    Built-in web search

    Not supported

    Log probabilities

    Not supported

    Deterministic seed

    Not supported

    Parallel tool calls

    Not supported

    Prompt caching

    Supported

    Strengths and limitations

    Strengths

    • Low input cost at $0.25 per million tokens suits high-volume workloads.
    • A 128K context window covers most focused application workflows.
    • Supports text & reasoning workloads in one model.

    Limitations

    • Real-world cost still depends on prompt length, response length and provider-specific billing rules.
    • The 128K context window is smaller than several long-context alternatives.
    • Arena Elo and MMLU-Pro are directional; test accuracy, latency and reliability on your own workload before committing.

    Arena Elo, MMLU-Pro and pricing figures are illustrative; validate current vendor terms before purchase.

    Same provider

    More models from Inception

    View provider →
    Mercury 2.5

    Inception

    260K

    $0.04 in · $0.15 out

    View model

    Frequently asked questions about Mercury 2

    Mercury 2 is a text & reasoning model from Inception. Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...