PriceIndex
NV

Nemotron 3 Super

nemotron-3-super-120b-a12b

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer... Built by NVIDIA.

Prices updated

Input price

$0.09

per 1M tokens · Standard

Output price

$0.40

per 1M tokens · Standard

Input limit

262K

tokens

Output limit

236K

tokens

Input formats

TextImageVideoAudioPDF

Output formats

TextImageVideoAudioPDF
Compare

Nemotron 3 Super price history

4 price records since 9/25/2026 · 2 changes

Nemotron 3 Super cost calculator

Input20M tokens
Output5M tokens

$3.70/month

Standard pricing

    Overview

    What is Nemotron 3 Super?

    Nemotron 3 Super is a text & reasoning model from NVIDIA. NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer... Its 262K context window and $0.09 input price make it a candidate for cost-sensitive, high-throughput applications.

    Content workflowsClassification

    Rates

    Nemotron 3 Super pricing

    Pricing is based on token usage. Provider-specific caching, batch, regional, and tool charges may affect the final cost.

    Input · Standard

    $0.09

    Per 1M tokens

    Cached input · Standard

    Not available

    No listed cached-input rate

    Capabilities

    Nemotron 3 Super Tools

    Tools available when using Nemotron 3 Super through supported provider APIs.

    Function calling

    Supported

    Structured outputs

    Supported

    JSON mode

    Supported

    Reasoning

    Supported

    Built-in web search

    Not supported

    Log probabilities

    Supported

    Deterministic seed

    Supported

    Parallel tool calls

    Not supported

    Prompt caching

    Not supported

    Strengths and limitations

    Strengths

    • Low input cost at $0.09 per million tokens suits high-volume workloads.
    • 262K context supports large documents, repositories and extended conversations.
    • Supports text & reasoning workloads in one model.

    Limitations

    • Generated tokens cost 5× more than input tokens, which matters for verbose responses.
    • Large context capacity does not guarantee consistent retrieval across the entire prompt.
    • Arena Elo and MMLU-Pro are directional; test accuracy, latency and reliability on your own workload before committing.

    Arena Elo, MMLU-Pro and pricing figures are illustrative; validate current vendor terms before purchase.

    Same provider

    More models from NVIDIA

    View provider →

    $0.20 in · $0.20 out

    View model

    $0.05 in · $0.14 out

    View model

    $0.06 in · $0.24 out

    View model

    $0.50 in · $2.20 out

    View model

    Frequently asked questions about Nemotron 3 Super

    Nemotron 3 Super is a text & reasoning model from NVIDIA. NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...