# DeepSeek V4 Pro API Pricing

As of October 10, 2026, DeepSeek V4 Pro costs $1.32 per 1M input tokens and $3.96 per 1M output tokens on DeepSeek, with cache reads at $0.044 and cache writes at $0 per 1M tokens. Across 17 providers on LiteLLM, the input price ranges from $0.435 to $2.40. It has a 1,000,000-token context window and returns up to 393,216 tokens per response.

- Page: https://models.litellm.ai/models/deepseek-v4-pro
- Lab: DeepSeek
- LiteLLM model name: `deepseek-v4-pro`
- Context window: 1,000,000 tokens
- Max output: 393,216 tokens
- Added to LiteLLM: June 2026
- Prices updated: October 10, 2026

## Pricing by provider

Token prices in USD per 1M tokens.

| Provider | Model name | Input | Output | Cache read | Per image | Per second | Context |
|---|---|---|---|---|---|---|---|
| DeepSeek | `deepseek-v4-pro` | $1.32 | $3.96 | $0.044 | – | – | 1M |
| DeepSeek | `deepseek/deepseek-v4-pro` | $1.32 | $3.96 | $0.044 | – | – | 1M |
| AIHubMix | `aihubmix/deepseek-v4-pro` | $1.69 | $3.38 | $0.1403 | – | – | 1M |
| Alibaba Cloud Model Studio | `dashscope/deepseek-v4-pro` | $2.40 | $4.80 | $0.20 | – | – | 1M |
| Azure AI Foundry | `azure_ai/deepseek-v4-pro` | $1.74 | $3.48 | $0.145 | – | – | 1M |
| Baseten | `baseten/deepseek-ai/DeepSeek-V4-Pro` | $1.74 | $3.48 | $0.145 | – | – | 1M |
| Baseten (snapshot 0813) | `baseten/deepseek-ai/DeepSeek-V4-Pro-0813` | $1.32 | $3.96 | $0.132 | – | – | 1M |
| Databricks (snapshot 0813) | `databricks/databricks-deepseek-v4-pro-0813` | $1.32 | $3.96 | $0.132 | – | – | 1M |
| DeepInfra | `deepinfra/deepseek-ai/DeepSeek-V4-Pro` | $1.30 | $2.60 | $0.10 | – | – | 1M |
| DeepInfra (snapshot 0813) | `deepinfra/deepseek-ai/DeepSeek-V4-Pro-0813` | $1.30 | $2.60 | $0.10 | – | – | 1M |
| Fireworks AI | `fireworks_ai/accounts/fireworks/models/deepseek-v4-pro` | $1.20 | $1.20 | $0.60 | – | – | 1M |
| Fireworks AI (snapshot 0813) | `fireworks_ai/accounts/fireworks/models/deepseek-v4-pro-0813` | $1.32 | $3.96 | $0.044 | – | – | 1M |
| Fireworks AI (snapshot 0813) | `fireworks_ai/deepseek-v4-pro-0813` | $1.32 | $3.96 | $0.044 | – | – | 1M |
| Nebius AI Studio | `nebius/deepseek-ai/DeepSeek-V4-Pro` | $1.75 | $3.50 | – | – | – | 1M |
| Nebius AI Studio (snapshot 0813) | `nebius/deepseek-ai/DeepSeek-V4-Pro-0813` | $1.32 | $3.96 | – | – | – | 979K |
| Novita AI | `novita/deepseek/deepseek-v4-pro` | $1.60 | $3.20 | $0.135 | – | – | 1M |
| Novita AI (snapshot 0813) | `novita/deepseek/deepseek-v4-pro-0813` | $1.32 | $3.96 | $0.132 | – | – | 1M |
| Perplexity (snapshot 0813) | `perplexity/perplexity/deepseek-v4-pro-0813` | $1.32 | $3.96 | $0.044 | – | – | – |
| Qianwen AI Platform | `qwen_ai_platform/deepseek-v4-pro` | $2.40 | $4.80 | $0.20 | – | – | 1M |
| QwenCloud | `qwencloud/deepseek-v4-pro` | $2.40 | $4.80 | $0.20 | – | – | 1M |
| Sail (snapshot 0813) | `sail/deepseek-ai/DeepSeek-V4-Pro-0813` | $0.92 | $2.77 | $0.04 | – | – | 1M |
| Tencent TokenHub | `tencent/deepseek-v4-pro` | $0.435 | $0.87 | $0.003625 | – | – | 1M |
| Together AI (snapshot 0813) | `together_ai/deepseek-ai/DeepSeek-V4-Pro-0813` | $1.32 | $3.96 | $0.13 | – | – | 1M |
| Weights & Biases | `wandb/deepseek-ai/DeepSeek-V4-Pro` | $1.15 | $2.55 | $0.20 | – | – | 1M |
| Weights & Biases (snapshot 0813) | `wandb/deepseek-ai/DeepSeek-V4-Pro-0813` | $1.31 | $3.96 | $0.044 | – | – | 1M |

## Cost examples

1,000 requests with 2,000 input and 500 output tokens each:

- DeepSeek: $4.62
- Tencent TokenHub: $1.305
- Sail: $3.225
- Weights & Biases: $3.575

100 long-document requests with 100,000 input and 2,000 output tokens each:

- DeepSeek: $13.99
- Tencent TokenHub: $4.524
- Sail: $9.754
- Weights & Biases: $12.01

## Features

Supported: Function calling, Parallel tool calls, Structured output, Reasoning, Prompt caching.
Not supported: Image input.

## Use it with LiteLLM

```python
from litellm import completion

response = completion(
    model="deepseek-v4-pro",
    messages=[{"role": "user", "content": "Hello!"}],
)
```

## FAQ

### How much does DeepSeek V4 Pro cost?

DeepSeek V4 Pro costs $1.32 per 1M input tokens and $3.96 per 1M output tokens on DeepSeek, with cache reads at $0.044 and cache writes at $0 per 1M tokens. Across 17 providers on LiteLLM, the input price ranges from $0.435 to $2.40.

### What is the context window of DeepSeek V4 Pro?

DeepSeek V4 Pro has a 1,000,000-token context window and returns up to 393,216 tokens per response.

### Which providers offer DeepSeek V4 Pro?

DeepSeek V4 Pro is available from 17 providers through LiteLLM: DeepSeek, AIHubMix, Alibaba Cloud Model Studio, Azure AI Foundry, Baseten, Databricks, DeepInfra, Fireworks AI, Nebius AI Studio, Novita AI, Perplexity and Qianwen AI Platform and 5 more. The lowest price is on Tencent TokenHub.

### How do I call DeepSeek V4 Pro with an OpenAI-compatible API?

Use the LiteLLM Python SDK with model="deepseek-v4-pro", or add the model to the LiteLLM proxy's config.yaml and send requests to /v1/chat/completions from any OpenAI client. LiteLLM tracks the cost of every request at the prices on this page.

### What features does DeepSeek V4 Pro support?

DeepSeek V4 Pro supports function calling, parallel tool calls, structured output, reasoning and prompt caching. It does not support image input.

### When did LiteLLM add DeepSeek V4 Pro?

DeepSeek V4 Pro was added to LiteLLM's model price file in June 2026.

Source: LiteLLM model price file, https://github.com/BerriAI/litellm/blob/main/model_prices_and_context_window.json
