PriceIndex

Models, measured side by side

DeepSeek V3 vs Phi-4

Compare DeepSeek V3, Phi-4 by token pricing, context length, benchmark results, speed and tool support.

Monthly workload

Monthly cost breakdown

InputOutput

DeepSeek V3

$10 / mo

Phi-4

$2.10 / mo

Performance comparison

DeepSeek V3

8.5

Phi-4

5.9

Overview

Released
DeepSeek V3
2024-12
Phi-4
2024-12
Status
DeepSeek V3
Active
Phi-4
Active
Input formats
DeepSeek V3
text
Phi-4
text
Output formats
DeepSeek V3
text
Phi-4
text

Pricing & capacity

Input / 1M tokens
DeepSeek V3
$0.26
Phi-4
$0.07
Output / 1M tokens
DeepSeek V3
$1.03
Phi-4
$0.14
Monthly cost
DeepSeek V3
$10
Phi-4
$2.10
Context window
DeepSeek V3
164K
Phi-4
16K
Maximum output
DeepSeek V3
16K
Phi-4
15K

Benchmarks & performance

Arena Elo
DeepSeek V3
1,211
Phi-4
1,123
MMLU-Pro
DeepSeek V3
75.4%
Phi-4
60.7%
Intelligence Index
DeepSeek V3
8.5
Phi-4
5.9
GPQA
DeepSeek V3
55.7%
Phi-4
57.5%
Humanity’s Last Exam
DeepSeek V3
2.9%
Phi-4
3.8%
SciCode
DeepSeek V3
35.8%
Phi-4
Not listed
Output speed
DeepSeek V3
Not listed
Phi-4
39 tokens/s
Time to first token
DeepSeek V3
Not listed
Phi-4
2.59 s

Tools & features

Function calling
DeepSeek V3
Supported
Phi-4
Not listed
Structured outputs
DeepSeek V3
Supported
Phi-4
Supported
JSON mode
DeepSeek V3
Supported
Phi-4
Supported
Deterministic seed
DeepSeek V3
Supported
Phi-4
Supported

API & availability

API identifier
DeepSeek V3
deepseek-v3
Phi-4
phi-4
API providers
DeepSeek V3
StreamLake, DeepInfra
Phi-4
DeepInfra
Knowledge cutoff
DeepSeek V3
2024-07-31
Phi-4
2024-06-30
Prices checked
DeepSeek V3
Oct 8, 2026
Phi-4
Oct 8, 2026

About the models

DeepSeek V3

MoE open model with frontier-class quality at low cost.

Full pricing & details

Phi-4

14B small language model, great for edge and cheap tasks.

Full pricing & details

Comparison FAQ

DeepSeek V3: $0.26 input and $1.03 output per million tokens. Phi-4: $0.07 input and $0.14 output per million tokens. The cheaper choice depends on your input-to-output ratio; a model can have a lower input rate but a higher output rate.

Popular model comparisons