Models, measured side by side
DeepSeek V4.1 Flash vs MiniMax M3
Compare DeepSeek V4.1 Flash, MiniMax M3 by token pricing, context length, benchmark results, speed and tool support.
Monthly workload
Monthly cost breakdown
InputOutput
DeepSeek V4.1 Flash
$7.00 / mo
MiniMax M3
$12 / mo
Performance comparison
DeepSeek V4.1 Flash
39.5
MiniMax M3
29.2
Overview
Pricing & capacity
| Pricing & capacity | DeepSeek V4.1 Flash | MiniMax M3 |
|---|---|---|
| Input / 1M tokens | $0.05 | $0.30 |
| Output / 1M tokens | $1.20 | $1.20 |
| Cached input / 1M | $0.02 | $0.06 |
| Monthly cost | $7.00 | $12 |
| Context window | 1M | 1M |
| Maximum output | 944K | 512K |
Input / 1M tokens
- DeepSeek V4.1 Flash
- $0.05
- MiniMax M3
- $0.30
Output / 1M tokens
- DeepSeek V4.1 Flash
- $1.20
- MiniMax M3
- $1.20
Cached input / 1M
- DeepSeek V4.1 Flash
- $0.02
- MiniMax M3
- $0.06
Monthly cost
- DeepSeek V4.1 Flash
- $7.00
- MiniMax M3
- $12
Context window
- DeepSeek V4.1 Flash
- 1M
- MiniMax M3
- 1M
Maximum output
- DeepSeek V4.1 Flash
- 944K
- MiniMax M3
- 512K
Benchmarks & performance
| Benchmarks & performance | DeepSeek V4.1 Flash | MiniMax M3 |
|---|---|---|
| Intelligence Index | 39.5 | 29.2 |
| GPQA | Not listed | 92.9% |
| Humanity’s Last Exam | 39.2% | 39% |
| SciCode | 51.9% | 47.1% |
| Output speed | 213 tokens/s | 97 tokens/s |
| Time to first token | 1.05 s | 2.03 s |
| DesignArena rating | 1,325 | 1,252 |
| DesignArena win rate | 52.1% | 50.9% |
Intelligence Index
- DeepSeek V4.1 Flash
- 39.5
- MiniMax M3
- 29.2
GPQA
- DeepSeek V4.1 Flash
- Not listed
- MiniMax M3
- 92.9%
Humanity’s Last Exam
- DeepSeek V4.1 Flash
- 39.2%
- MiniMax M3
- 39%
SciCode
- DeepSeek V4.1 Flash
- 51.9%
- MiniMax M3
- 47.1%
Output speed
- DeepSeek V4.1 Flash
- 213 tokens/s
- MiniMax M3
- 97 tokens/s
Time to first token
- DeepSeek V4.1 Flash
- 1.05 s
- MiniMax M3
- 2.03 s
DesignArena rating
- DeepSeek V4.1 Flash
- 1,325
- MiniMax M3
- 1,252
DesignArena win rate
- DeepSeek V4.1 Flash
- 52.1%
- MiniMax M3
- 50.9%
Tools & features
| Tools & features | DeepSeek V4.1 Flash | MiniMax M3 |
|---|---|---|
| Function calling | Supported | Supported |
| Structured outputs | Supported | Supported |
| JSON mode | Supported | Supported |
| Reasoning | Supported | Supported |
| Log probabilities | Supported | Supported |
| Deterministic seed | Supported | Supported |
| Prompt caching | Supported | Supported |
Function calling
- DeepSeek V4.1 Flash
- Supported
- MiniMax M3
- Supported
Structured outputs
- DeepSeek V4.1 Flash
- Supported
- MiniMax M3
- Supported
JSON mode
- DeepSeek V4.1 Flash
- Supported
- MiniMax M3
- Supported
Reasoning
- DeepSeek V4.1 Flash
- Supported
- MiniMax M3
- Supported
Log probabilities
- DeepSeek V4.1 Flash
- Supported
- MiniMax M3
- Supported
Deterministic seed
- DeepSeek V4.1 Flash
- Supported
- MiniMax M3
- Supported
Prompt caching
- DeepSeek V4.1 Flash
- Supported
- MiniMax M3
- Supported
API & availability
| API & availability | DeepSeek V4.1 Flash | MiniMax M3 |
|---|---|---|
| API identifier | deepseek-v4.1-flash | minimax-m3 |
| Prices checked | Oct 8, 2026 | Oct 8, 2026 |
API identifier
- DeepSeek V4.1 Flash
deepseek-v4.1-flash- MiniMax M3
minimax-m3
Prices checked
- DeepSeek V4.1 Flash
- Oct 8, 2026
- MiniMax M3
- Oct 8, 2026
About the models
DeepSeek V4.1 Flash
DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...
Full pricing & detailsMiniMax M3
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
Full pricing & detailsComparison FAQ
DeepSeek V4.1 Flash: $0.05 input and $1.20 output per million tokens. MiniMax M3: $0.30 input and $1.20 output per million tokens. The cheaper choice depends on your input-to-output ratio; a model can have a lower input rate but a higher output rate.
