LiteLLMModels

PPLX Embed Context V1 4B API Pricing

As of October 10, 2026, PPLX Embed Context V1 4B costs $0.05 per 1M input tokens on Perplexity. It has a 32,768-token context window.

Model ID perplexity/pplx-embed-context-v1-4bPrices updated Checked
Input
$0.05 / 1M tokens
Output
$0 / 1M tokens
Context window
32,768 tokens
Providers
1
Added to LiteLLM
August 20, 2026

PPLX Embed Context V1 4B pricing by provider

1 listing across 1 provider. Token prices are in US dollars per 1M tokens. The highlighted row is the price quoted above.

ProviderModel name in LiteLLMInputOutputContext
Perplexityperplexity/pplx-embed-context-v1-4b$0.05$032K

What PPLX Embed Context V1 4B costs in practice

Embedding 10,000 documents of 500 tokens each

Perplexity$0.25

Use PPLX Embed Context V1 4B with LiteLLM

Call PPLX Embed Context V1 4B through the LiteLLM Python SDK, or put it behind the LiteLLM proxy and send OpenAI-format requests to /v1/embeddings. LiteLLM tracks the cost of every request at these prices.

Python SDK
from litellm import embedding

response = embedding(
    model="perplexity/pplx-embed-context-v1-4b",
    input=["hello world"],
)
Proxy config.yaml
model_list:
  - model_name: pplx-embed-context-v1-4b
    litellm_params:
      model: perplexity/pplx-embed-context-v1-4b
      api_key: os.environ/PERPLEXITYAI_API_KEY
Call the proxy
curl http://0.0.0.0:4000/v1/embeddings \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer sk-1234" \
  -d '{
    "model": "pplx-embed-context-v1-4b",
    "input": "hello world"
  }'

PPLX Embed Context V1 4B FAQ

How much does PPLX Embed Context V1 4B cost?

PPLX Embed Context V1 4B costs $0.05 per 1M input tokens on Perplexity.

What is the context window of PPLX Embed Context V1 4B?

PPLX Embed Context V1 4B has a 32,768-token context window.

Which providers offer PPLX Embed Context V1 4B?

PPLX Embed Context V1 4B is available from Perplexity through LiteLLM, using the model name perplexity/pplx-embed-context-v1-4b.

How do I call PPLX Embed Context V1 4B with an OpenAI-compatible API?

Use the LiteLLM Python SDK with model="perplexity/pplx-embed-context-v1-4b", or add the model to the LiteLLM proxy's config.yaml and send requests to /v1/embeddings from any OpenAI client. LiteLLM tracks the cost of every request at the prices on this page.

When did LiteLLM add PPLX Embed Context V1 4B?

PPLX Embed Context V1 4B was added to LiteLLM's model price file on August 20, 2026.

Sources

Prices come from LiteLLM's open-source model_prices_and_context_window.json, which LiteLLM uses to calculate the cost of every request. Provider pricing pages:

LiteLLM docs: Perplexity.

Provider pages: Perplexity.