Skip to main content
Perplexity

Search documentation

Type to search this documentation.

On this pageOverview

Perplexity with LiteLLM

LiteLLM is a Python SDK and proxy server that gives you a single OpenAI-compatible interface to 100+ LLM providers. Perplexity's Agent API — with third-party models like GPT-5, Claude, and Gemini routed through Perplexity — is a first-class provider in LiteLLM.

Bash
pip install litellm

LiteLLM reads your Perplexity API key from the environment:

Bash
export PERPLEXITY_API_KEY="your_api_key_here"

Get API Key

Generate your Perplexity API key from the API portal.

Use litellm.responses to call the Agent API, which routes through Perplexity to third-party models with tool orchestration and presets.

Python
from litellm import responses
import os

os.environ["PERPLEXITY_API_KEY"] = "your_api_key_here"

response = responses(
    model="perplexity/preset/low",
    input="What are the latest developments in AI?",
    custom_llm_provider="perplexity",
)

print(response.output)

Available presets: fast, low, medium, high, xhigh.

Python
from litellm import responses

response = responses(
    model="perplexity/openai/gpt-5.6-sol",
    input="Research quantum computing breakthroughs and cite sources.",
    custom_llm_provider="perplexity",
    tools=[
        {"type": "web_search"},
        {"type": "fetch_url"},
    ],
    instructions="Use web_search and fetch_url to gather citations.",
    max_output_tokens=1000,
    temperature=0.7,
)

print(response.output)
Python
from litellm import responses

response = responses(
    model="perplexity/preset/low",
    input="Extract key facts about the Eiffel Tower.",
    custom_llm_provider="perplexity",
    text={
        "format": {
            "type": "json_schema",
            "name": "facts",
            "schema": {
                "type": "object",
                "properties": {
                    "name": {"type": "string"},
                    "height_meters": {"type": "number"},
                    "year_built": {"type": "integer"},
                },
                "required": ["name", "height_meters", "year_built"],
            },
            "strict": True,
        }
    },
)

Prefix any Agent API model ID with perplexity/ (for example, perplexity/openai/gpt-5.6-sol). See the Agent API model list for the canonical, up-to-date catalogue.

Run LiteLLM as a self-hosted proxy that fronts Perplexity (and any other provider) behind a single OpenAI-compatible endpoint.

YAML
model_list:
  - model_name: perplexity-low
    litellm_params:
      model: perplexity/preset/low
      api_key: os.environ/PERPLEXITY_API_KEY
Bash
litellm --config /path/to/config.yaml
Bash
curl http://0.0.0.0:4000/v1/responses \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer anything" \
  -d '{
    "model": "perplexity-low",
    "input": "What are the latest developments in AI?",
    "tools": [{"type": "web_search"}]
  }'

Need help with the integration?

Suggest an edit

Propose a replacement for this page. The site team reviews it before applying any changes.

Export
Documentation menu