Sonar Pro API pricing
Published on-demand rates for Perplexity's perplexity/sonar-pro, captured 2026-09-21 from public pricing data and committed to our archive. Context window 200k tokens; first observed in our archive 2025-02-17.
| Meter | USD per 1M tokens |
|---|---|
| Input | $3.00 |
| Output | $15.00 |
| Cached inputno published cache-read rate; the whole input bills at the standard rate | — |
Standard on-demand rates. Excludes batch endpoints and negotiated pricing; confirm on Perplexity's pricing page before committing.
What it costs on a real workload
Per 1M tokens processed, at the rates above. How the four profiles are defined.
| Profile | In / out mix | Cached share | Cost |
|---|---|---|---|
| Retrieval | 800k / 200k | 70% | $5.40 |
| Chat | 500k / 500k | 30% | $9.00 |
| Content | 200k / 800k | 20% | $12.60 |
| Agent | 900k / 100k | 85% | $4.20 |
Observed rate history
No repricing observed since this model entered our archive on 2025-02-17. When a rate moves, the change appears here and on the sitewide ticker.
What the record shows
Sonar Pro has not repriced once since it entered our archive on 2025-02-17. That is the ordinary case rather than the exception: providers ship new models at lower rates far more often than they cut the price of one already in service, which is why a cheaper option usually arrives as a new name rather than as a smaller number on this page.
Against the 40 models we price, it is the 37th cheapest for retrieval-shaped work, which reads a great deal and writes little, and the 36th cheapest for content-shaped work, which does the reverse. The two are close, so it holds its position whichever way your workload leans.
It charges 5.0 times more for output than for input, against a median of 5.0 times across the set. That is close to typical for the models we track.
Inside Perplexity's own lineup it is the most expensive of the 4 models we track on this workload. That is worth settling before comparing it against another vendor's flagship, because the cheaper alternatives are on the same account, behind the same key, with no migration to do.
30 of the 40 models we price offer a larger context window, and 1 matches it exactly. Context is a capability limit rather than a cost one: paying more buys no extra room unless the larger window is what you are paying for.
Perplexity publishes no cache-read rate for it, so a prompt prefix you send on every call costs the same as text the model has never seen. On a cache-heavy workload that is a structural disadvantage against models that publish one, and it is why the agent profile above prices its whole input at the standard rate.
Whether it accepts image input is not stated in the source we capture, and we do not guess at it.
Its 200k token context window holds roughly 150,000 words of English text in a single request. A workload that does not fit cannot run here at any price, which is a capability limit rather than a cost one.
Priced near this one
Closest to Sonar Pro on a retrieval workload, cheaper and more expensive alike. Cost only; we publish no quality score.
| Model | Provider | Input | Output | vs Sonar Pro |
|---|---|---|---|---|
| GPT 5.6 | OpenAI | $4.00 | $20.00 | 4% less |
| Claude Opus 5 | Anthropic | $5.00 | $25.00 | 20% more |
| GPT 5.3 | OpenAI | $1.75 | $14.00 | 39% less |
| Sonar Reasoning Pro | Perplexity | $2.00 | $8.00 | 41% less |
Retrieval profile: 800k in, 200k out, 70% cached where offered.
No affiliate or referral arrangements: the console link carries no parameters and nothing on this page is paid placement. Sonar Pro is a trademark of its owner; LLMPrice.io is independent and unaffiliated.