Mai Code 1.1 Flash API Pricing
Mai Code 1.1 Flash is available through LiteLLM. It has a 128,000-token context window and returns up to 128,000 tokens per response.
- Context window
- 128,000 tokens
- Max output
- 128,000 tokens
- Providers
- 1
Mai Code 1.1 Flash pricing by provider
1 listing across 1 provider. Token prices are in US dollars per 1M tokens. The highlighted row is the price quoted above.
| Provider | Model name in LiteLLM | Context |
|---|---|---|
| GitHub Copilot | github_copilot/mai-code-1.1-flash | 128K |
Mai Code 1.1 Flash features
- Function calling
- Parallel tool calls
- Structured output
- Image input
Use Mai Code 1.1 Flash with LiteLLM
Call Mai Code 1.1 Flash through the LiteLLM Python SDK, or put it behind the LiteLLM proxy and send OpenAI-format requests to /v1/responses. LiteLLM tracks the cost of every request at these prices.
from litellm import completion
response = completion(
model="github_copilot/mai-code-1.1-flash",
messages=[{"role": "user", "content": "Hello!"}],
)curl http://0.0.0.0:4000/v1/chat/completions \
-H "Content-Type: application/json" \
-H "Authorization: Bearer sk-1234" \
-d '{
"model": "mai-code-1.1-flash",
"messages": [{"role": "user", "content": "Hello!"}]
}'Mai Code 1.1 Flash FAQ
What is the context window of Mai Code 1.1 Flash?
Mai Code 1.1 Flash has a 128,000-token context window and returns up to 128,000 tokens per response.
What features does Mai Code 1.1 Flash support?
Mai Code 1.1 Flash supports function calling, parallel tool calls, structured output and image input.
Sources
Prices come from LiteLLM's open-source model_prices_and_context_window.json, which LiteLLM uses to calculate the cost of every request.
LiteLLM docs: GitHub Copilot.
Provider pages: GitHub Copilot.
