Text Embedding 3 Small Inference API Pricing
Text Embedding 3 Small Inference is available through LiteLLM. It has an 8,191-token context window.
- Context window
- 8,191 tokens
- Providers
- 1
- Added to LiteLLM
- December 2025
Text Embedding 3 Small Inference pricing by provider
1 listing across 1 provider. Token prices are in US dollars per 1M tokens. The highlighted row is the price quoted above.
| Provider | Model name in LiteLLM | Context |
|---|---|---|
| GitHub Copilot | github_copilot/text-embedding-3-small-inference | 8K |
Use Text Embedding 3 Small Inference with LiteLLM
Call Text Embedding 3 Small Inference through the LiteLLM Python SDK, or put it behind the LiteLLM proxy and send OpenAI-format requests to /v1/embeddings. LiteLLM tracks the cost of every request at these prices.
from litellm import embedding
response = embedding(
model="github_copilot/text-embedding-3-small-inference",
input=["hello world"],
)curl http://0.0.0.0:4000/v1/embeddings \
-H "Content-Type: application/json" \
-H "Authorization: Bearer sk-1234" \
-d '{
"model": "text-embedding-3-small-inference",
"input": "hello world"
}'Text Embedding 3 Small Inference FAQ
What is the context window of Text Embedding 3 Small Inference?
Text Embedding 3 Small Inference has an 8,191-token context window.
When did LiteLLM add Text Embedding 3 Small Inference?
Text Embedding 3 Small Inference was added to LiteLLM's model price file in December 2025.
Sources
Prices come from LiteLLM's open-source model_prices_and_context_window.json, which LiteLLM uses to calculate the cost of every request.
LiteLLM docs: GitHub Copilot.
Provider pages: GitHub Copilot.
