Perplexity Models
All 4 Perplexity Models
Open in model list| llama-3.1-sonar-small-128k-online | $0.30$0.30/M |
| llama-3.1-sonar-large-128k-online | $1.20$1.20/M |
| sonar | $1.60$1.60/M |
| llama-3.1-sonar-huge-128k-online | $5.60$5.60/M |
Perplexity on AIHubMix
Which Perplexity model should I start with?
llama-3.1-sonar-small-128k-online at $0.30/M input — the cheapest entry here that declares a token price. Move up to llama-3.1-sonar-huge-128k-online when answer quality matters more than cost.
Why are there several entries for the same model?
Because each row is a route you can call, not a model release. Some IDs name an upstream (azure-, alicloud-, cc-), and some differ only in capitalisation, kept so older integrations keep working.
The catalog does not carry a field saying which of those a given row is, so this page does not sort them into buckets it would have to invent. Every row shows that route’s own price, context and speed — compare those directly, and open a model to see the upstreams that serve it.
Do I need a separate Perplexity account?
No. One AIHubMix key covers every model on this page, and switching between them is a change to the model string — billing, rate limits, and logs stay in one place.
Start calling Perplexity in one line
One key, one endpoint, 844 models across 35 providers.
