MiMo-V2.6-Flash
mimo-v2.6-flashMiMo-V2.6-Flash is an open-source foundation model developed by Xiaomi. Built on a Mixture-of-Experts architecture with 309B total parameters and 15B activated per token, it employs a hybrid attention mechanism for... Built by Xiaomi.
Prices updated
Input price
$0.14
per 1M tokens · Standard
Output price
$0.28
per 1M tokens · Standard
Input limit
1.1M
tokens
Output limit
131K
tokens
Input formats
Output formats
MiMo-V2.6-Flash price history
2 price records since 9/25/2026
MiMo-V2.6-Flash cost calculator
$4.20/month
Standard pricing
Overview
What is MiMo-V2.6-Flash?
MiMo-V2.6-Flash is a text & reasoning and vision model from Xiaomi. MiMo-V2.6-Flash is an open-source foundation model developed by Xiaomi. Built on a Mixture-of-Experts architecture with 309B total parameters and 15B activated per token, it employs a hybrid attention mechanism for... Its 1.1M context window and $0.14 input price make it a candidate for cost-sensitive, high-throughput applications.
Benchmarks
MiMo-V2.6-Flash benchmarks & speed
How MiMo-V2.6-Flash scores on standardized evaluations, and where it lands among every model we track.
37.9
Intelligence Index
56tok/s
Output speed
Reasoning
HLE
35.1%
Humanity's Last Exam · top 30%
Coding
SciCode
51.3%
Python for scientific computing · top 44%
Latency & design
- Time to first token
- 5.15s
Independent scores from Artificial Analysis and DesignArena · updated 10/5/2026. Higher is better; ranks compare against every model we track with that score.
Rates
MiMo-V2.6-Flash pricing
Pricing is based on token usage. Provider-specific caching, batch, regional, and tool charges may affect the final cost.
Input · Standard
$0.14
Per 1M tokens
Cached input · Standard
$0.0028
Per 1M tokens
Capabilities
MiMo-V2.6-Flash Tools
Tools available when using MiMo-V2.6-Flash through supported provider APIs.
Function calling
Supported
Structured outputs
Supported
JSON mode
Supported
Reasoning
Supported
Built-in web search
Not supported
Log probabilities
Supported
Deterministic seed
Supported
Parallel tool calls
Not supported
Prompt caching
Supported
Strengths and limitations
Strengths
- Low input cost at $0.14 per million tokens suits high-volume workloads.
- 1.1M context supports large documents, repositories and extended conversations.
- Supports text & reasoning and vision workloads in one model.
Limitations
- Real-world cost still depends on prompt length, response length and provider-specific billing rules.
- Large context capacity does not guarantee consistent retrieval across the entire prompt.
- Arena Elo and MMLU-Pro are directional; test accuracy, latency and reliability on your own workload before committing.
Arena Elo, MMLU-Pro and pricing figures are illustrative; validate current vendor terms before purchase.
Same provider
