Hybrid reasoning Claude model with strong agentic coding. Built by Anthropic.
Prices updated
Input price
$3.00
per 1M tokens · Standard
Output price
$15
per 1M tokens · Standard
Input limit
200K
tokens
Output limit
64K
tokens
Input formats
Output formats
Claude Sonnet 4 price history
12 price records since 4/1/2026
Overview
What is Claude Sonnet 4?
Claude Sonnet 4 is a text & reasoning and vision and code & slms model from Anthropic. Hybrid reasoning Claude model with strong agentic coding. Its 200K context window and $3.00 input price make it a candidate for quality-focused production applications.
Benchmarks
Claude Sonnet 4 benchmarks & speed
How Claude Sonnet 4 scores on standardized evaluations, and where it lands among every model we track.
16.6
Intelligence Index
1145
DesignArena Elo
Reasoning
GPQA Diamond
68.3%
Graduate-level scientific reasoning · top 72%
HLE
4.3%
Humanity's Last Exam · top 85%
Latency & design
- DesignArena win rate
- 53.4%
- Design battles judged
- 17,542
Independent scores from Artificial Analysis and DesignArena · updated 10/8/2026. Higher is better; ranks compare against every model we track with that score.
Rates
Claude Sonnet 4 pricing
Pricing is based on token usage. Provider-specific caching, batch, regional, and tool charges may affect the final cost.
Input · Standard
$3.00
Per 1M tokens
Cached input · Standard
$0.30
Per 1M tokens
Capabilities
Claude Sonnet 4 Tools
Tools available when using Claude Sonnet 4 through supported provider APIs.
Function calling
Supported
Structured outputs
Not supported
JSON mode
Not supported
Reasoning
Supported
Built-in web search
Not supported
Log probabilities
Not supported
Deterministic seed
Not supported
Parallel tool calls
Not supported
Prompt caching
Supported
Where to run it
Claude Sonnet 4 API providers
2 providers serve Claude Sonnet 4. Prices are per 1M tokens; uptime is the last 24 hours.
| Provider | Input | Output | Cached | Context | Uptime |
|---|---|---|---|---|---|
| Amazon Bedrock | $3.00 | $15 | $0.30 | 200K | 100% |
| Amazon Bedrock | $3.00 | $15 | $0.30 | 200K | 100% |
- Input
- $3.00
- Output
- $15
- Cached
- $0.30
- Context
- 200K
- Uptime
- 100%
- Input
- $3.00
- Output
- $15
- Cached
- $0.30
- Context
- 200K
- Uptime
- 100%
Strengths and limitations
Strengths
- Measured results (AA Intelligence 16.6 · DesignArena Elo 1145 · Arena Elo 1250 · MMLU-Pro 81.9%) position it for demanding production work.
- 200K context supports large documents, repositories and extended conversations.
- Supports text & reasoning and vision and code & slms workloads in one model.
Limitations
- Generated tokens cost 5× more than input tokens, which matters for verbose responses.
- Large context capacity does not guarantee consistent retrieval across the entire prompt.
- Arena Elo and MMLU-Pro are directional; test accuracy, latency and reliability on your own workload before committing.
Arena Elo, MMLU-Pro and pricing figures are illustrative; validate current vendor terms before purchase.
Same provider
More models from Anthropic
Peer set
