Models, measured side by side
DeepSeek V4 Flash 0423 vs gpt-oss-20b
Compare DeepSeek V4 Flash 0423, gpt-oss-20b by token pricing, context length, benchmark results, speed and tool support.
Monthly workload
Monthly cost breakdown
InputOutput
DeepSeek V4 Flash 0423
$0.88 / mo
gpt-oss-20b
$0.81 / mo
Performance comparison
DeepSeek V4 Flash 0423
34.3
gpt-oss-20b
9
Overview
Pricing & capacity
| Pricing & capacity | DeepSeek V4 Flash 0423 | gpt-oss-20b |
|---|---|---|
| Input / 1M tokens | $0.03 | $0.02 |
| Output / 1M tokens | $0.06 | $0.09 |
| Cached input / 1M | $0.0075 | $0.0090 |
| Monthly cost | $0.88 | $0.81 |
| Context window | 1M | 131K |
| Maximum output | 944K | 33K |
Input / 1M tokens
- DeepSeek V4 Flash 0423
- $0.03
- gpt-oss-20b
- $0.02
Output / 1M tokens
- DeepSeek V4 Flash 0423
- $0.06
- gpt-oss-20b
- $0.09
Cached input / 1M
- DeepSeek V4 Flash 0423
- $0.0075
- gpt-oss-20b
- $0.0090
Monthly cost
- DeepSeek V4 Flash 0423
- $0.88
- gpt-oss-20b
- $0.81
Context window
- DeepSeek V4 Flash 0423
- 1M
- gpt-oss-20b
- 131K
Maximum output
- DeepSeek V4 Flash 0423
- 944K
- gpt-oss-20b
- 33K
Benchmarks & performance
| Benchmarks & performance | DeepSeek V4 Flash 0423 | gpt-oss-20b |
|---|---|---|
| Intelligence Index | 34.3 | 9 |
| GPQA | 90.8% | 68.8% |
| Humanity’s Last Exam | 38.6% | 11% |
| SciCode | 50.3% | 38.9% |
| Output speed | 215 tokens/s | 207 tokens/s |
| Time to first token | 0.95 s | 0.85 s |
| DesignArena rating | 1,209 | Not listed |
| DesignArena win rate | 48.9% | Not listed |
Intelligence Index
- DeepSeek V4 Flash 0423
- 34.3
- gpt-oss-20b
- 9
GPQA
- DeepSeek V4 Flash 0423
- 90.8%
- gpt-oss-20b
- 68.8%
Humanity’s Last Exam
- DeepSeek V4 Flash 0423
- 38.6%
- gpt-oss-20b
- 11%
SciCode
- DeepSeek V4 Flash 0423
- 50.3%
- gpt-oss-20b
- 38.9%
Output speed
- DeepSeek V4 Flash 0423
- 215 tokens/s
- gpt-oss-20b
- 207 tokens/s
Time to first token
- DeepSeek V4 Flash 0423
- 0.95 s
- gpt-oss-20b
- 0.85 s
DesignArena rating
- DeepSeek V4 Flash 0423
- 1,209
- gpt-oss-20b
- Not listed
DesignArena win rate
- DeepSeek V4 Flash 0423
- 48.9%
- gpt-oss-20b
- Not listed
Tools & features
| Tools & features | DeepSeek V4 Flash 0423 | gpt-oss-20b |
|---|---|---|
| Function calling | Supported | Supported |
| Structured outputs | Supported | Supported |
| JSON mode | Supported | Supported |
| Reasoning | Supported | Supported |
| Log probabilities | Supported | Supported |
| Deterministic seed | Supported | Supported |
| Prompt caching | Supported | Supported |
Function calling
- DeepSeek V4 Flash 0423
- Supported
- gpt-oss-20b
- Supported
Structured outputs
- DeepSeek V4 Flash 0423
- Supported
- gpt-oss-20b
- Supported
JSON mode
- DeepSeek V4 Flash 0423
- Supported
- gpt-oss-20b
- Supported
Reasoning
- DeepSeek V4 Flash 0423
- Supported
- gpt-oss-20b
- Supported
Log probabilities
- DeepSeek V4 Flash 0423
- Supported
- gpt-oss-20b
- Supported
Deterministic seed
- DeepSeek V4 Flash 0423
- Supported
- gpt-oss-20b
- Supported
Prompt caching
- DeepSeek V4 Flash 0423
- Supported
- gpt-oss-20b
- Supported
API & availability
| API & availability | DeepSeek V4 Flash 0423 | gpt-oss-20b |
|---|---|---|
| API identifier | deepseek-v4-flash | gpt-oss-20b |
| Knowledge cutoff | Not listed | 2024-06-30 |
| Prices checked | Oct 8, 2026 | Oct 8, 2026 |
API identifier
- DeepSeek V4 Flash 0423
deepseek-v4-flash- gpt-oss-20b
gpt-oss-20b
Knowledge cutoff
- DeepSeek V4 Flash 0423
- Not listed
- gpt-oss-20b
- 2024-06-30
Prices checked
- DeepSeek V4 Flash 0423
- Oct 8, 2026
- gpt-oss-20b
- Oct 8, 2026
About the models
DeepSeek V4 Flash 0423
DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...
Full pricing & detailsgpt-oss-20b
gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...
Full pricing & detailsComparison FAQ
DeepSeek V4 Flash 0423: $0.03 input and $0.06 output per million tokens. gpt-oss-20b: $0.02 input and $0.09 output per million tokens. The cheaper choice depends on your input-to-output ratio; a model can have a lower input rate but a higher output rate.
