PriceIndex

Models, measured side by side

DeepSeek V4.1 Flash vs Hy3

Compare DeepSeek V4.1 Flash, Hy3 by token pricing, context length, benchmark results, speed and tool support.

Monthly workload

Monthly cost breakdown

InputOutput

DeepSeek V4.1 Flash

$7.00 / mo

Hy3

$5.28 / mo

Performance comparison

DeepSeek V4.1 Flash

39.5

Hy3

25.3

Overview

Released
DeepSeek V4.1 Flash
2026-09-10
Hy3
2026-07-06
Status
DeepSeek V4.1 Flash
Active
Hy3
Active
Input formats
DeepSeek V4.1 Flash
textimage
Hy3
text
Output formats
DeepSeek V4.1 Flash
text
Hy3
text

Pricing & capacity

Input / 1M tokens
DeepSeek V4.1 Flash
$0.05
Hy3
$0.13
Output / 1M tokens
DeepSeek V4.1 Flash
$1.20
Hy3
$0.53
Cached input / 1M
DeepSeek V4.1 Flash
$0.02
Hy3
$0.03
Monthly cost
DeepSeek V4.1 Flash
$7.00
Hy3
$5.28
Context window
DeepSeek V4.1 Flash
1M
Hy3
262K
Maximum output
DeepSeek V4.1 Flash
944K
Hy3
128K

Benchmarks & performance

Intelligence Index
DeepSeek V4.1 Flash
39.5
Hy3
25.3
GPQA
DeepSeek V4.1 Flash
Not listed
Hy3
89.7%
Humanity’s Last Exam
DeepSeek V4.1 Flash
39.2%
Hy3
33.5%
SciCode
DeepSeek V4.1 Flash
51.9%
Hy3
48.6%
Output speed
DeepSeek V4.1 Flash
213 tokens/s
Hy3
88 tokens/s
Time to first token
DeepSeek V4.1 Flash
1.05 s
Hy3
3.08 s
DesignArena rating
DeepSeek V4.1 Flash
1,325
Hy3
1,180
DesignArena win rate
DeepSeek V4.1 Flash
52.1%
Hy3
41.1%

Tools & features

Function calling
DeepSeek V4.1 Flash
Supported
Hy3
Supported
Structured outputs
DeepSeek V4.1 Flash
Supported
Hy3
Supported
JSON mode
DeepSeek V4.1 Flash
Supported
Hy3
Supported
Reasoning
DeepSeek V4.1 Flash
Supported
Hy3
Supported
Log probabilities
DeepSeek V4.1 Flash
Supported
Hy3
Not listed
Deterministic seed
DeepSeek V4.1 Flash
Supported
Hy3
Supported
Prompt caching
DeepSeek V4.1 Flash
Supported
Hy3
Supported

API & availability

API identifier
DeepSeek V4.1 Flash
deepseek-v4.1-flash
Hy3
hy3
Prices checked
DeepSeek V4.1 Flash
Oct 8, 2026
Hy3
Oct 8, 2026

About the models

DeepSeek V4.1 Flash

DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...

Full pricing & details

Hy3

Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B active, 192 experts with top-8 routing) built for reasoning, agentic workflows, and real-world production use. It supports a configurable reasoning effort:...

Full pricing & details

Comparison FAQ

DeepSeek V4.1 Flash: $0.05 input and $1.20 output per million tokens. Hy3: $0.13 input and $0.53 output per million tokens. The cheaper choice depends on your input-to-output ratio; a model can have a lower input rate but a higher output rate.

Popular model comparisons