Provider

Perplexity models through odnoga

Models that answer from the live web rather than from training data alone. Use them where the answer has to be current, and price them as what they are: a search and a generation in one call.

Models
2
Cheapest input / 1M
$1
Largest context
200K
With cached input
0/2

When freshness is the requirement

Every other model on the gateway answers from what it learned during training, with a cutoff. These do not — which makes them the right choice for questions about prices, availability, news and anything else that changed this week, and the wrong choice for deterministic work where you want the same answer every time.

Budget them separately

A web-grounded answer does more work than a plain completion, and the cost reflects it. If you expose this to end users, give it its own budget and its own per-user cap rather than pooling it with your cheap traffic — odnoga enforces caps before the spend happens, not after the invoice.

Your plan

Vendor cost is the list price. Your price is that plus your plan margin — the same figures that appear on your invoice.

Every model here, cheapest first

ModelVendorContextIn / 1MOut / 1MCached inYour price in
Sonar
sonar
Perplexity127K$1$1$1.08Details
Sonar Pro
sonar-pro
Perplexity200K$3$15$3.23Details

Your price column is the vendor cost +7%.

One key, every model on this page.

Change the model with a routing rule instead of a deploy, and bill every call to the customer who made it.