Models, measured side by side
DeepSeek V4.1 Flash vs Hy3
Compare DeepSeek V4.1 Flash, Hy3 by token pricing, context length, benchmark results, speed and tool support.
Monthly workload
Monthly cost breakdown
InputOutput
DeepSeek V4.1 Flash
$7.00 / mo
Hy3
$5.28 / mo
Performance comparison
DeepSeek V4.1 Flash
39.5
Hy3
25.3
Overview
Pricing & capacity
| Pricing & capacity | DeepSeek V4.1 Flash | Hy3 |
|---|---|---|
| Input / 1M tokens | $0.05 | $0.13 |
| Output / 1M tokens | $1.20 | $0.53 |
| Cached input / 1M | $0.02 | $0.03 |
| Monthly cost | $7.00 | $5.28 |
| Context window | 1M | 262K |
| Maximum output | 944K | 128K |
Input / 1M tokens
- DeepSeek V4.1 Flash
- $0.05
- Hy3
- $0.13
Output / 1M tokens
- DeepSeek V4.1 Flash
- $1.20
- Hy3
- $0.53
Cached input / 1M
- DeepSeek V4.1 Flash
- $0.02
- Hy3
- $0.03
Monthly cost
- DeepSeek V4.1 Flash
- $7.00
- Hy3
- $5.28
Context window
- DeepSeek V4.1 Flash
- 1M
- Hy3
- 262K
Maximum output
- DeepSeek V4.1 Flash
- 944K
- Hy3
- 128K
Benchmarks & performance
| Benchmarks & performance | DeepSeek V4.1 Flash | Hy3 |
|---|---|---|
| Intelligence Index | 39.5 | 25.3 |
| GPQA | Not listed | 89.7% |
| Humanity’s Last Exam | 39.2% | 33.5% |
| SciCode | 51.9% | 48.6% |
| Output speed | 213 tokens/s | 88 tokens/s |
| Time to first token | 1.05 s | 3.08 s |
| DesignArena rating | 1,325 | 1,180 |
| DesignArena win rate | 52.1% | 41.1% |
Intelligence Index
- DeepSeek V4.1 Flash
- 39.5
- Hy3
- 25.3
GPQA
- DeepSeek V4.1 Flash
- Not listed
- Hy3
- 89.7%
Humanity’s Last Exam
- DeepSeek V4.1 Flash
- 39.2%
- Hy3
- 33.5%
SciCode
- DeepSeek V4.1 Flash
- 51.9%
- Hy3
- 48.6%
Output speed
- DeepSeek V4.1 Flash
- 213 tokens/s
- Hy3
- 88 tokens/s
Time to first token
- DeepSeek V4.1 Flash
- 1.05 s
- Hy3
- 3.08 s
DesignArena rating
- DeepSeek V4.1 Flash
- 1,325
- Hy3
- 1,180
DesignArena win rate
- DeepSeek V4.1 Flash
- 52.1%
- Hy3
- 41.1%
Tools & features
| Tools & features | DeepSeek V4.1 Flash | Hy3 |
|---|---|---|
| Function calling | Supported | Supported |
| Structured outputs | Supported | Supported |
| JSON mode | Supported | Supported |
| Reasoning | Supported | Supported |
| Log probabilities | Supported | Not listed |
| Deterministic seed | Supported | Supported |
| Prompt caching | Supported | Supported |
Function calling
- DeepSeek V4.1 Flash
- Supported
- Hy3
- Supported
Structured outputs
- DeepSeek V4.1 Flash
- Supported
- Hy3
- Supported
JSON mode
- DeepSeek V4.1 Flash
- Supported
- Hy3
- Supported
Reasoning
- DeepSeek V4.1 Flash
- Supported
- Hy3
- Supported
Log probabilities
- DeepSeek V4.1 Flash
- Supported
- Hy3
- Not listed
Deterministic seed
- DeepSeek V4.1 Flash
- Supported
- Hy3
- Supported
Prompt caching
- DeepSeek V4.1 Flash
- Supported
- Hy3
- Supported
API & availability
| API & availability | DeepSeek V4.1 Flash | Hy3 |
|---|---|---|
| API identifier | deepseek-v4.1-flash | hy3 |
| Prices checked | Oct 8, 2026 | Oct 8, 2026 |
API identifier
- DeepSeek V4.1 Flash
deepseek-v4.1-flash- Hy3
hy3
Prices checked
- DeepSeek V4.1 Flash
- Oct 8, 2026
- Hy3
- Oct 8, 2026
About the models
DeepSeek V4.1 Flash
DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...
Full pricing & detailsHy3
Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B active, 192 experts with top-8 routing) built for reasoning, agentic workflows, and real-world production use. It supports a configurable reasoning effort:...
Full pricing & detailsComparison FAQ
DeepSeek V4.1 Flash: $0.05 input and $1.20 output per million tokens. Hy3: $0.13 input and $0.53 output per million tokens. The cheaper choice depends on your input-to-output ratio; a model can have a lower input rate but a higher output rate.
