GPT-4 O Preview API Pricing
GPT-4 O Preview is available through LiteLLM. It has a 64,000-token context window and returns up to 4,096 tokens per response.
- Context window
- 64,000 tokens
- Max output
- 4,096 tokens
- Providers
- 1
- Added to LiteLLM
- December 2025
GPT-4 O Preview pricing by provider
1 listing across 1 provider. Token prices are in US dollars per 1M tokens. The highlighted row is the price quoted above.
| Provider | Model name in LiteLLM | Context |
|---|---|---|
| GitHub Copilot | github_copilot/gpt-4-o-preview | 64K |
GPT-4 O Preview features
- Function calling
- Parallel tool calls
Use GPT-4 O Preview with LiteLLM
Call GPT-4 O Preview through the LiteLLM Python SDK, or put it behind the LiteLLM proxy and send OpenAI-format requests to /v1/chat/completions. LiteLLM tracks the cost of every request at these prices.
from litellm import completion
response = completion(
model="github_copilot/gpt-4-o-preview",
messages=[{"role": "user", "content": "Hello!"}],
)curl http://0.0.0.0:4000/v1/chat/completions \
-H "Content-Type: application/json" \
-H "Authorization: Bearer sk-1234" \
-d '{
"model": "gpt-4-o-preview",
"messages": [{"role": "user", "content": "Hello!"}]
}'GPT-4 O Preview FAQ
What is the context window of GPT-4 O Preview?
GPT-4 O Preview has a 64,000-token context window and returns up to 4,096 tokens per response.
What features does GPT-4 O Preview support?
GPT-4 O Preview supports function calling and parallel tool calls.
When did LiteLLM add GPT-4 O Preview?
GPT-4 O Preview was added to LiteLLM's model price file in December 2025.
Sources
Prices come from LiteLLM's open-source model_prices_and_context_window.json, which LiteLLM uses to calculate the cost of every request.
LiteLLM docs: GitHub Copilot.
Provider pages: GitHub Copilot.
