Router Models & Pricing
Overview
Section titled “Overview”Every model below is available through both Router endpoints — Chat Completions and Messages — under the same creator/model-name id. Prices are in USD per 1 million tokens, and you are billed at the requested model's rates regardless of how the request is served.
For the live catalog, query GET /models; the tables below show the current catalog.
Perplexity-Hosted Models
Section titled “Perplexity-Hosted Models”Open-source models hosted by Perplexity — the perplexity/ prefix reflects who serves the model, not who created it. The catalog includes models from Moonshot AI, NVIDIA, and Z.AI:
| Model | Input ($/1M) | Output ($/1M) | Cache read ($/1M) | Docs |
|---|---|---|---|---|
perplexity/kimi-k3 |
3.00 | 15.00 | 0.30 | Kimi K3 |
perplexity/glm-5.3 |
1.40 | 4.40 | 0.26 | GLM |
perplexity/glm-5.3-flash |
0.15 | 0.50 | 0.03 | GLM-5.3 Flash |
perplexity/nemotron-3-ultra-550b-a55b |
0.25 | 2.50 | 0.25 | Nemotron 3 Ultra |
Listing Models Programmatically
Section titled “Listing Models Programmatically”curl 'https://api.perplexity.ai/router/v1/models' \
-H "Authorization: Bearer $PERPLEXITY_API_KEY" | jqThe response lists every available model sorted by id, including each model's base token prices. Requesting a model that is not in the catalog returns a 400 naming the invalid model — the catalog is also the allowlist.