PriceIndex

Models, measured side by side

GPT-4.1 vs GPT-4o

Compare GPT-4.1, GPT-4o by token pricing, context length, benchmark results, speed and tool support.

Monthly workload

Monthly cost breakdown

InputOutput

GPT-4.1

$80 / mo

GPT-4o

$100 / mo

Performance comparison

GPT-4.1

12.7

GPT-4o

8.4

Overview

Released
GPT-4.1
2025-04
GPT-4o
2024-05
Status
GPT-4.1
Active
GPT-4o
Active
Input formats
GPT-4.1
imagetextfile
GPT-4o
textimagefile
Output formats
GPT-4.1
text
GPT-4o
text

Pricing & capacity

Input / 1M tokens
GPT-4.1
$2.00
GPT-4o
$2.50
Output / 1M tokens
GPT-4.1
$8.00
GPT-4o
$10
Cached input / 1M
GPT-4.1
$0.50
GPT-4o
$1.25
Monthly cost
GPT-4.1
$80
GPT-4o
$100
Context window
GPT-4.1
1M
GPT-4o
128K
Maximum output
GPT-4.1
33K
GPT-4o
16K

Benchmarks & performance

Arena Elo
GPT-4.1
1,200
GPT-4o
1,178
MMLU-Pro
GPT-4.1
73.6%
GPT-4o
69.9%
Intelligence Index
GPT-4.1
12.7
GPT-4o
8.4
GPQA
GPT-4.1
66.6%
GPT-4o
54.3%
Humanity’s Last Exam
GPT-4.1
4.2%
GPT-4o
2.4%
SciCode
GPT-4.1
Not listed
GPT-4o
Not listed

6 of 10

Tools & features

Function calling
GPT-4.1
Supported
GPT-4o
Supported
Structured outputs
GPT-4.1
Supported
GPT-4o
Supported
JSON mode
GPT-4.1
Supported
GPT-4o
Supported
Reasoning
GPT-4.1
Not listed
GPT-4o
Not listed
Built-in web search
GPT-4.1
Not listed
GPT-4o
Supported
Log probabilities
GPT-4.1
Not listed
GPT-4o
Supported

6 of 9

API & availability

API identifier
GPT-4.1
gpt-4.1
GPT-4o
gpt-4o
API providers
GPT-4.1
Azure, OpenAI, Azure
GPT-4o
Azure, OpenAI
Knowledge cutoff
GPT-4.1
2024-06-30
GPT-4o
2023-10-31
Prices checked
GPT-4.1
Oct 8, 2026
GPT-4o
Oct 8, 2026

About the models

GPT-4.1

Non-reasoning GPT-4.1 with a 1M-token context window, strong at instruction following and coding.

Full pricing & details

GPT-4o

Multimodal GPT-4o "omni" model for text and image input.

Full pricing & details

Comparison FAQ

GPT-4.1: $2.00 input and $8.00 output per million tokens. GPT-4o: $2.50 input and $10 output per million tokens. The cheaper choice depends on your input-to-output ratio; a model can have a lower input rate but a higher output rate.

Popular model comparisons