Sonar Reasoning Pro
Language Search-grounded Available Full comparison ↗ perplexity/sonar-reasoning-pro · by Perplexity
· decoder-only-transformer
Search-grounded model. Unlike a standard base model, this offering combines an underlying LLM with live web retrieval and citations — not a directly comparable like-for-like with a non-search-grounded entry. Billed with the token pricing below plus a per-request search-context fee (see Pricing).
Pricing — 1 offering(s)
Input tokens
- $2.00 / 1M tokens (input) Current 2026-07-26 → present
Output tokens
- $8.00 / 1M tokens (output) Current 2026-07-26 → present
Search context fee (billed on top of token pricing, per query)
- low $6.00 / 1k requests
- medium $10.00 / 1k requests
- high $14.00 / 1k requests
Showing the active price and any recorded history. Full pricing history is available via the paid API — see API docs.
See how Sonar Reasoning Pro fits into a cost-aware routing setup
See how →Capability profile
Operator guidance
Route here when the task genuinely needs chain-of-thought analysis on top of cited web search and can tolerate higher latency. For simple factual lookups it is slower and dearer than Sonar with no benefit. For exhaustive hundreds-of-sources reports, Sonar Deep Research is purpose-built and prices its citation/reasoning tokens separately. Outside Perplexity, pairing a standalone reasoning model (o-series, DeepSeek R1, Claude with extended thinking) with your own retrieval gives more control but loses the tuned grounding.
Use cases
- Complex, multi-step analytical questions that need both live web evidence and visible step-by-step reasoning
- Tasks with strict formatting or instruction constraints where a plain search model drifts
- Comparative / evaluative research questions (surfacing more citations per search than the plain search models) that don't need a full Deep Research report
Limitations
- Search-grounded only, with a mandatory per-request search-context fee ($6-$14 / 1K requests) on top of token cost
- Perplexity explicitly does not recommend it for simple factual queries, basic retrieval, or speed-critical applications
- Chain-of-thought tokens are billed as normal output tokens on this variant, so verbose reasoning directly inflates cost
- Underlying base model is not named in current Perplexity docs; the DeepSeek-R1 basis is a prior Perplexity claim
- Legacy Sonar Chat Completions surface migrating to the Agent API (support until 2026-09-27, Perplexity changelog)
- Capability ratings are qualitative from Perplexity's positioning; no independent published benchmark for the API model