Perplexity: Sonar Pro
Perplexity Sonar Pro API pricing
Sonar Pro costs $1.20 per million input tokens and $6.00 per million output tokens on Oxyy. These are the pay-as-you-go prices after the current 60% discount (Launch Promo; standard price $3.00 input and $15.00 output). Web search is billed on top, at $7.20 per 1K searches. It accepts text, images and files and returns text.
Advanced search-grounded model with deeper citation metadata and better multi-hop query handling.
Providers
Where this model runs, and what each route costs. Requests are sent to a healthy provider automatically; if one errors, the gateway retries against another serving the same model.
| Provider | Input /M | Output /M | Cache read /M |
|---|---|---|---|
| Perplexity | $1.20 | $6.00 | — |
Capabilities (9)
Supported parameters (19)
Any other parameter you send is ignored rather than rejected.
Pricing
What this model costs on Oxyy, pay as you go. "Charged" is the rate a request is billed at; the standard price is shown beside it for reference.
| Rate | Standard price | Charged | Unit |
|---|---|---|---|
| Input | $3.00 | $1.20 | per 1M input tokens |
| Output | $15.00 | $6.00 | per 1M output tokens |
| Web search | $18.00 | $7.20 | per 1K searches |
Discount. Charged rates include the current "Launch Promo" discount of 60% on every rate.
Specifications
- Model ID
sonar-pro- Developer
- Perplexity
- Type
- Text
- Input
- Text, images and files
- Output
- Text
- Series
- Sonar
- Image input
- Up to 20 MB each
- File input
- Up to 100 MB each
- Endpoint
/v1/chat/completions- Status
- Deprecated (Sep 27, 2026)
Best for
Quickstart
Call sonar-pro with a POST to /v1/chat/completions on https://api.oxyy.ai, using your Oxyy API key from the OXYY_API_KEY environment variable. The same key works for every model in the catalog.
import os
from openai import OpenAI
client = OpenAI(
base_url="https://api.oxyy.ai/v1",
api_key=os.environ["OXYY_API_KEY"],
)
response = client.chat.completions.create(
model="sonar-pro",
messages=[
{"role": "user", "content": "Explain unified AI APIs in one sentence."},
],
)
print(response.choices[0].message.content)Availability
The share of requests to this model that completed successfully over the last 30 days. When an upstream provider errors, the gateway retries against another provider serving the same model, so a single provider incident does not necessarily show up here.
Activity
Token volume and request traffic to this model over time. Daily totals, UTC.
Prompt tokens measure input size. Reasoning tokens show internal thinking before a response. Completion tokens reflect total output length. Reasoning is not priced separately for this model, so it is not itemised here.
Frequently asked questions
What is Sonar Pro?
How much does Sonar Pro cost?
Does Sonar Pro support tool calling and structured outputs?
What inputs and outputs does Sonar Pro support?
Which API endpoint does Sonar Pro use?
Alternatives to Sonar Pro
Models of the same kind from other vendors, closest in price first. Prices are pay-as-you-go on Oxyy, through the same API key.
