Provider
Perplexity API pricing: every model compared
Perplexity offers 5 models through its API in our data, priced from $1 to $3 per million input tokens. The newest is Sonar Pro Search, first listed on 2025-10-30.
- Models
- 5
- Cheapest input
- $1
- Priciest input
- $3
- Largest context
- 200,000
- With caching
- none
- Open weights
- none
Cheapest
SonarPerplexity$1 in · $1 outLargest context
Sonar ProPerplexity200,000 tokensNewest
Sonar Pro SearchPerplexity2025-10-30Line-up
Every Perplexity model and price
| Model | Input / 1M | Output / 1M | Context | Released |
|---|---|---|---|---|
| Sonar Pro Search | $3 | $15 | 200,000 | 2025-10-30 |
| Sonar Deep Research | $2 | $8 | 128,000 | 2025-03-07 |
| Sonar Pro | $3 | $15 | 200,000 | 2025-03-07 |
| Sonar Reasoning Pro | $2 | $8 | 128,000 | 2025-03-07 |
| Sonar | $1 | $1 | 127,072 | 2025-01-27 |
Prices in US dollars per 1M tokens, newest first. Release dates are when our source first listed the model.
Overview
Perplexity’s pricing at a glance
The spread between Perplexity’s cheapest and most expensive model is large: Sonar Pro Search costs 3× more per input token than Sonar. Picking the smallest model that does the job well is usually the biggest saving available, ahead of any discount.
None of the line-up has a published cached-input price in our data. No batch prices are published in our data. 4 of 5 accept images as input.
FAQ
Perplexity API questions
How much does the Perplexity API cost?
Perplexity's 5 models range from $1 to $3 per million input tokens, and output costs more than input on almost every model. The exact bill depends on your token volumes, which you can estimate in the LLM cost calculator.
What is Perplexity's cheapest model?
By blended price (three parts input to one part output) it is Sonar, at $1 input and $1 output per million tokens. The most expensive is Sonar Pro Search.
Which Perplexity model has the largest context window?
Sonar Pro, with 200,000 tokens per request.
Does Perplexity offer prompt caching or batch discounts?
In our data, none of Perplexity's models have a published cached-input price and none have a published batch price. Both can cut costs substantially for repeated prompts or work that can wait.