Models, measured side by side
GPT-4.1 vs GPT-4o
Compare GPT-4.1, GPT-4o by token pricing, context length, benchmark results, speed and tool support.
Monthly workload
Monthly cost breakdown
InputOutput
GPT-4.1
$80 / mo
GPT-4o
$100 / mo
Performance comparison
GPT-4.1
12.7
GPT-4o
8.4
Overview
Pricing & capacity
| Pricing & capacity | GPT-4.1 | GPT-4o |
|---|---|---|
| Input / 1M tokens | $2.00 | $2.50 |
| Output / 1M tokens | $8.00 | $10 |
| Cached input / 1M | $0.50 | $1.25 |
| Monthly cost | $80 | $100 |
| Context window | 1M | 128K |
| Maximum output | 33K | 16K |
Input / 1M tokens
- GPT-4.1
- $2.00
- GPT-4o
- $2.50
Output / 1M tokens
- GPT-4.1
- $8.00
- GPT-4o
- $10
Cached input / 1M
- GPT-4.1
- $0.50
- GPT-4o
- $1.25
Monthly cost
- GPT-4.1
- $80
- GPT-4o
- $100
Context window
- GPT-4.1
- 1M
- GPT-4o
- 128K
Maximum output
- GPT-4.1
- 33K
- GPT-4o
- 16K
Benchmarks & performance
| Benchmarks & performance | GPT-4.1 | GPT-4o |
|---|---|---|
| Arena Elo | 1,200 | 1,178 |
| MMLU-Pro | 73.6% | 69.9% |
| Intelligence Index | 12.7 | 8.4 |
| GPQA | 66.6% | 54.3% |
| Humanity’s Last Exam | 4.2% | 2.4% |
| SciCode | Not listed | Not listed |
| Output speed | 166 tokens/s | 145 tokens/s |
| Time to first token | 0.83 s | 0.84 s |
| DesignArena rating | 1,029 | 864 |
| DesignArena win rate | 50.9% | 34.8% |
Arena Elo
- GPT-4.1
- 1,200
- GPT-4o
- 1,178
MMLU-Pro
- GPT-4.1
- 73.6%
- GPT-4o
- 69.9%
Intelligence Index
- GPT-4.1
- 12.7
- GPT-4o
- 8.4
GPQA
- GPT-4.1
- 66.6%
- GPT-4o
- 54.3%
Humanity’s Last Exam
- GPT-4.1
- 4.2%
- GPT-4o
- 2.4%
SciCode
- GPT-4.1
- Not listed
- GPT-4o
- Not listed
Output speed
- GPT-4.1
- 166 tokens/s
- GPT-4o
- 145 tokens/s
Time to first token
- GPT-4.1
- 0.83 s
- GPT-4o
- 0.84 s
DesignArena rating
- GPT-4.1
- 1,029
- GPT-4o
- 864
DesignArena win rate
- GPT-4.1
- 50.9%
- GPT-4o
- 34.8%
6 of 10
Tools & features
| Tools & features | GPT-4.1 | GPT-4o |
|---|---|---|
| Function calling | Supported | Supported |
| Structured outputs | Supported | Supported |
| JSON mode | Supported | Supported |
| Reasoning | Not listed | Not listed |
| Built-in web search | Not listed | Supported |
| Log probabilities | Not listed | Supported |
| Deterministic seed | Supported | Supported |
| Parallel tool calls | Not listed | Not listed |
| Prompt caching | Supported | Supported |
Function calling
- GPT-4.1
- Supported
- GPT-4o
- Supported
Structured outputs
- GPT-4.1
- Supported
- GPT-4o
- Supported
JSON mode
- GPT-4.1
- Supported
- GPT-4o
- Supported
Reasoning
- GPT-4.1
- Not listed
- GPT-4o
- Not listed
Built-in web search
- GPT-4.1
- Not listed
- GPT-4o
- Supported
Log probabilities
- GPT-4.1
- Not listed
- GPT-4o
- Supported
Deterministic seed
- GPT-4.1
- Supported
- GPT-4o
- Supported
Parallel tool calls
- GPT-4.1
- Not listed
- GPT-4o
- Not listed
Prompt caching
- GPT-4.1
- Supported
- GPT-4o
- Supported
6 of 9
API & availability
| API & availability | GPT-4.1 | GPT-4o |
|---|---|---|
| API identifier | gpt-4.1 | gpt-4o |
| API providers | Azure, OpenAI, Azure | Azure, OpenAI |
| Knowledge cutoff | 2024-06-30 | 2023-10-31 |
| Prices checked | Oct 8, 2026 | Oct 8, 2026 |
API identifier
- GPT-4.1
gpt-4.1- GPT-4o
gpt-4o
API providers
- GPT-4.1
- Azure, OpenAI, Azure
- GPT-4o
- Azure, OpenAI
Knowledge cutoff
- GPT-4.1
- 2024-06-30
- GPT-4o
- 2023-10-31
Prices checked
- GPT-4.1
- Oct 8, 2026
- GPT-4o
- Oct 8, 2026
About the models
GPT-4.1
Non-reasoning GPT-4.1 with a 1M-token context window, strong at instruction following and coding.
Full pricing & detailsComparison FAQ
GPT-4.1: $2.00 input and $8.00 output per million tokens. GPT-4o: $2.50 input and $10 output per million tokens. The cheaper choice depends on your input-to-output ratio; a model can have a lower input rate but a higher output rate.
