Skip to main content
Perplexity

Search documentation

Type to search this documentation.

On this pageOverview

Router Models & Pricing

Every model below is available through both Router endpoints — Chat Completions and Messages — under the same creator/model-name id. Prices are in USD per 1 million tokens, and you are billed at the requested model's rates regardless of how the request is served.

For the live catalog, query GET /models; the tables below show the current catalog.

Open-source models hosted by Perplexity — the perplexity/ prefix reflects who serves the model, not who created it. The catalog includes models from Moonshot AI, NVIDIA, and Z.AI:

Model Input ($/1M) Output ($/1M) Cache read ($/1M) Docs
perplexity/kimi-k3 3.00 15.00 0.30 Kimi K3
perplexity/glm-5.3 1.40 4.40 0.26 GLM
perplexity/glm-5.3-flash 0.15 0.50 0.03 GLM-5.3 Flash
perplexity/nemotron-3-ultra-550b-a55b 0.25 2.50 0.25 Nemotron 3 Ultra
Bash
curl 'https://api.perplexity.ai/router/v1/models' \
  -H "Authorization: Bearer $PERPLEXITY_API_KEY" | jq

The response lists every available model sorted by id, including each model's base token prices. Requesting a model that is not in the catalog returns a 400 naming the invalid model — the catalog is also the allowlist.

Suggest an edit

Propose a replacement for this page. The site team reviews it before applying any changes.

Export
Documentation menu