Sonar
sonarSonar is lightweight, affordable, fast, and simple to use — now featuring citations and the ability to customize sources. It is designed for companies seeking to integrate lightweight question-and-answer features... Built by Perplexity.
Prices updated
Input price
$1.00
per 1M tokens · Standard
Output price
$1.00
per 1M tokens · Standard
Input limit
127K
tokens
Output limit
114K
tokens
Input formats
Output formats
Sonar price history
2 price records since 9/25/2026
Sonar cost calculator
$25/month
Standard pricing
Overview
What is Sonar?
Sonar is a text & reasoning and vision model from Perplexity. Sonar is lightweight, affordable, fast, and simple to use — now featuring citations and the ability to customize sources. It is designed for companies seeking to integrate lightweight question-and-answer features... Its 127K context window and $1.00 input price make it a candidate for quality-focused production applications.
Benchmarks
Sonar benchmarks & speed
How Sonar scores on standardized evaluations, and where it lands among every model we track.
7.7
Intelligence Index
Reasoning
GPQA Diamond
47.1%
Graduate-level scientific reasoning · top 90%
HLE
4.9%
Humanity's Last Exam · top 81%
Independent scores from Artificial Analysis and DesignArena · updated 10/5/2026. Higher is better; ranks compare against every model we track with that score.
Rates
Sonar pricing
Pricing is based on token usage. Provider-specific caching, batch, regional, and tool charges may affect the final cost.
Input · Standard
$1.00
Per 1M tokens
Cached input · Standard
Not available
No listed cached-input rate
Capabilities
Sonar Tools
Tools available when using Sonar through supported provider APIs.
Function calling
Not supported
Structured outputs
Not supported
JSON mode
Not supported
Reasoning
Not supported
Built-in web search
Supported
Log probabilities
Not supported
Deterministic seed
Not supported
Parallel tool calls
Not supported
Prompt caching
Not supported
Strengths and limitations
Strengths
- Measured results (AA Intelligence 7.7) position it for demanding production work.
- A 127K context window covers most focused application workflows.
- Supports text & reasoning and vision workloads in one model.
Limitations
- Real-world cost still depends on prompt length, response length and provider-specific billing rules.
- The 127K context window is smaller than several long-context alternatives.
- Arena Elo and MMLU-Pro are directional; test accuracy, latency and reliability on your own workload before committing.
Arena Elo, MMLU-Pro and pricing figures are illustrative; validate current vendor terms before purchase.
Same provider
