Mistral Large API Pricing
As of October 10, 2026, Mistral Large costs $0.50 per 1M input tokens and $1.50 per 1M output tokens on Mistral AI, with cache reads at $0.05 per 1M tokens. Across 8 providers on LiteLLM, the input price ranges from $0.50 to $8.00. It has a 262,144-token context window and returns up to 262,144 tokens per response.
- Input
- $0.50 / 1M tokens
- Output
- $1.50 / 1M tokens
- Cache read
- $0.05 / 1M tokens
- Context window
- 262,144 tokens
- Max output
- 262,144 tokens
- Providers
- 8
- Added to LiteLLM
- February 2024
Mistral Large pricing by provider
18 listings across 8 providers. Token prices are in US dollars per 1M tokens. The highlighted row is the price quoted above.
| Provider | Model name in LiteLLM | Input | Output | Cache read | Context |
|---|---|---|---|---|---|
| Mistral AIlatest alias | mistral/mistral-large-latest | $0.50 | $1.50 | $0.05 | 256K |
| Amazon Bedrocksnapshot 2402 | mistral.mistral-large-2402-v1:0 | $4.00 | $12.00 | – | 32K |
| Amazon Bedrocksnapshot 2407 | mistral.mistral-large-2407-v1:0 | $2.00 | $6.00 | – | 128K |
| Azure AI Foundry | azure_ai/mistral-large | $4.00 | $12.00 | – | 32K |
| Azure AI Foundrysnapshot 2407 | azure_ai/mistral-large-2407 | $2.00 | $6.00 | – | 128K |
| Azure AI Foundrylatest alias | azure_ai/mistral-large-latest | $2.00 | $6.00 | – | 128K |
| Azure OpenAIsnapshot 2402 | azure/mistral-large-2402 | $8.00 | $24.00 | – | 32K |
| Azure OpenAIlatest alias | azure/mistral-large-latest | $8.00 | $24.00 | – | 32K |
| Google Vertex AIsnapshot 2411 | vertex_ai/mistral-large-2411 | $2.00 | $6.00 | – | 128K |
| Google Vertex AI | vertex_ai/mistral-large@2407 | $2.00 | $6.00 | – | 128K |
| Google Vertex AI | vertex_ai/mistral-large@2411-001 | $2.00 | $6.00 | – | 128K |
| Google Vertex AI | vertex_ai/mistral-large@latest | $2.00 | $6.00 | – | 128K |
| IBM watsonx.ai | watsonx/mistralai/mistral-large | $3.00 | $10.00 | – | 128K |
| Snowflake Cortex | snowflake/mistral-large | – | – | – | 32K |
| Vercel AI Gateway | vercel_ai_gateway/mistral/mistral-large | $2.00 | $6.00 | – | 32K |
| Amazon Bedrockeu-west-3, snapshot 2402 | bedrock/eu-west-3/mistral.mistral-large-2402-v1:0 | $5.20 | $15.60 | – | 32K |
| Amazon Bedrockus-east-1, snapshot 2402 | bedrock/us-east-1/mistral.mistral-large-2402-v1:0 | $4.00 | $12.00 | – | 32K |
| Amazon Bedrockus-west-2, snapshot 2402 | bedrock/us-west-2/mistral.mistral-large-2402-v1:0 | $4.00 | $12.00 | – | 32K |
What Mistral Large costs in practice
1,000 requests with 2,000 input and 500 output tokens each
| Mistral AI | $1.75 |
| Amazon Bedrock | $7.00 |
| Azure AI Foundry | $7.00 |
| Google Vertex AI | $7.00 |
100 long-document requests with 100,000 input and 2,000 output tokens each
| Mistral AI | $5.30 |
| Amazon Bedrock | $21.20 |
| Azure AI Foundry | $21.20 |
| Google Vertex AI | $21.20 |
Mistral Large features
- Function calling
- Structured output
- Image input
- Prompt caching
Use Mistral Large with LiteLLM
Call Mistral Large through the LiteLLM Python SDK, or put it behind the LiteLLM proxy and send OpenAI-format requests to /v1/chat/completions. LiteLLM tracks the cost of every request at these prices.
from litellm import completion
response = completion(
model="mistral/mistral-large-latest",
messages=[{"role": "user", "content": "Hello!"}],
)model_list:
- model_name: mistral-large
litellm_params:
model: mistral/mistral-large-latest
api_key: os.environ/MISTRAL_API_KEYcurl http://0.0.0.0:4000/v1/chat/completions \
-H "Content-Type: application/json" \
-H "Authorization: Bearer sk-1234" \
-d '{
"model": "mistral-large",
"messages": [{"role": "user", "content": "Hello!"}]
}'Mistral Large FAQ
How much does Mistral Large cost?
Mistral Large costs $0.50 per 1M input tokens and $1.50 per 1M output tokens on Mistral AI, with cache reads at $0.05 per 1M tokens. Across 8 providers on LiteLLM, the input price ranges from $0.50 to $8.00.
What is the context window of Mistral Large?
Mistral Large has a 262,144-token context window and returns up to 262,144 tokens per response.
Which providers offer Mistral Large?
Mistral Large is available from 8 providers through LiteLLM: Mistral AI, Amazon Bedrock, Azure AI Foundry, Azure OpenAI, Google Vertex AI, IBM watsonx.ai, Snowflake Cortex and Vercel AI Gateway. The lowest price is on Mistral AI.
How do I call Mistral Large with an OpenAI-compatible API?
Use the LiteLLM Python SDK with model="mistral/mistral-large-latest", or add the model to the LiteLLM proxy's config.yaml and send requests to /v1/chat/completions from any OpenAI client. LiteLLM tracks the cost of every request at the prices on this page.
What features does Mistral Large support?
Mistral Large supports function calling, structured output, image input and prompt caching.
When did LiteLLM add Mistral Large?
Mistral Large was added to LiteLLM's model price file in February 2024.
Sources
Prices come from LiteLLM's open-source model_prices_and_context_window.json, which LiteLLM uses to calculate the cost of every request. Provider pricing pages:
- docs.mistral.ai/models/mistral-large-3-25-12
- aws.amazon.com/bedrock/pricing/
- b0.p.awsstatic.com/pricing/2.0/meteredUnitMaps/bedrock/USD/current/bedrock.json
- learn.microsoft.com/en-us/azure/foundry/openai/concepts/retired-models
LiteLLM docs: Mistral AI, Amazon Bedrock, Azure AI Foundry, Azure OpenAI, Google Vertex AI, IBM watsonx.ai.
Provider pages: Mistral AI, Amazon Bedrock, Azure AI Foundry, Azure OpenAI, Google Vertex AI, IBM watsonx.ai, Snowflake Cortex, Vercel AI Gateway.
