Skip to content
llmwaves

Sonar Pro

Perplexity · released Mar 7, 2025

ProprietaryVision

Note: Sonar Pro pricing includes Perplexity search pricing. See [details here](https://docs.perplexity.ai/guides/pricing#detailed-pricing-breakdown-for-sonar-reasoning-pro-and-sonar-pro) For enterprises seeking more advanced capabilities, the Sonar Pro API can handle in-depth, multi-step queries with added extensibility, like...

Specification

Context window
200K
Max output
8K
Knowledge cutoff
Not stated
Parameters
Undisclosed
Licence
Proprietary
Serving providers
1
Moderated
No
Uptime
100.0%

Intelligence

9.1

33th percentile

Coding

Coding Index

Agentic

Agentic Index

Output speed

83 t/s

Median across providers

Latency

1.81s

Time to first token

Cost per task

$0.53

Estimated

Benchmarks

Where the score comes from

The Intelligence Index is a composite. These are the underlying evaluations this model was actually measured on.

Evaluation scores

Percentage correct · higher is better

  • MMLU-Pro
    75.5%
  • GPQA Diamond
    57.8%
  • AIME 2025
    29.0%
  • LiveCodeBench
    27.5%
  • SciCode
    22.6%
  • Humanity's Last Exam
    6.6%

An evaluation missing from this list was not run for this model — it is not a zero.

View as table
EvaluationScore
MMLU-Pro75.5%
GPQA Diamond57.8%
AIME 202529.0%
LiveCodeBench27.5%
SciCode22.6%
Humanity's Last Exam6.6%

Against its peers

Intelligence Index · this model highlighted, nearest peers in grey

Peers are the models sitting closest on the Intelligence Index — the set you would realistically choose between.

View as table
ModelIntelligence
Qwen3 VL 30B A3B Instruct9.9
R1 Distill Llama 70B9.8
GPT-4.1 Nano9.6
Sonar9.4
Qwen2.5 72B Instruct9.4
GPT-4o (2024-08-06)9.4
Sonar Pro9.1
GPT-4o (2024-05-13)8.4

Percentile among all indexed models

Intelligence33th

Pricing

What it costs to run

List prices per million tokens, plus what one representative task works out to.

List price

Input / 1M tokens
$3
Output / 1M tokens
$15
Cached input / 1M
Not offered
Blended 3:1
$6

One task, estimated

$0.53

Input tokens
50,000
Output tokens
25,000
Profile
Standard

Estimated from list pricing: 50K input tokens plus 80K output tokens for reasoning models (25K for non-reasoning).

Serving providers

Speed and latency figures are medians across these providers, so a widely-served model reports a blend rather than any single endpoint.