The smallest model in the Ministral 3 family, Ministral 3 3B is a powerful, efficient tiny language model with vision capabilities. Built by Mistral AI.
Prices updated
Input price
$0.10
per 1M tokens · Standard
Output price
$0.10
per 1M tokens · Standard
Input limit
131K
tokens
Output limit
105K
tokens
Input formats
Output formats
Ministral 3 3B 2512 price history
2 price records since 9/25/2026
Ministral 3 3B 2512 cost calculator
$2.50/month
Standard pricing
Overview
What is Ministral 3 3B 2512?
Ministral 3 3B 2512 is a text & reasoning and vision model from Mistral AI. The smallest model in the Ministral 3 family, Ministral 3 3B is a powerful, efficient tiny language model with vision capabilities. Its 131K context window and $0.10 input price make it a candidate for cost-sensitive, high-throughput applications.
Benchmarks
Ministral 3 3B 2512 benchmarks & speed
How Ministral 3 3B 2512 scores on standardized evaluations, and where it lands among every model we track.
1014
DesignArena Elo
Latency & design
- DesignArena win rate
- 37.3%
- Design battles judged
- 2,852
Independent scores from Artificial Analysis and DesignArena · updated 10/8/2026. Higher is better; ranks compare against every model we track with that score.
Rates
Ministral 3 3B 2512 pricing
Pricing is based on token usage. Provider-specific caching, batch, regional, and tool charges may affect the final cost.
Input · Standard
$0.10
Per 1M tokens
Cached input · Standard
$0.01
Per 1M tokens
Capabilities
Ministral 3 3B 2512 Tools
Tools available when using Ministral 3 3B 2512 through supported provider APIs.
Function calling
Supported
Structured outputs
Supported
JSON mode
Supported
Reasoning
Not supported
Built-in web search
Not supported
Log probabilities
Not supported
Deterministic seed
Supported
Parallel tool calls
Not supported
Prompt caching
Supported
Strengths and limitations
Strengths
- Low input cost at $0.10 per million tokens suits high-volume workloads.
- A 131K context window covers most focused application workflows.
- Supports text & reasoning and vision workloads in one model.
Limitations
- Real-world cost still depends on prompt length, response length and provider-specific billing rules.
- The 131K context window is smaller than several long-context alternatives.
- Arena Elo and MMLU-Pro are directional; test accuracy, latency and reliability on your own workload before committing.
Arena Elo, MMLU-Pro and pricing figures are illustrative; validate current vendor terms before purchase.
Same provider
