LiteLLMModels

Mai Code 1.1 Flash API Pricing

Mai Code 1.1 Flash is available through LiteLLM. It has a 128,000-token context window and returns up to 128,000 tokens per response.

Model ID github_copilot/mai-code-1.1-flashPrices updated Checked
Context window
128,000 tokens
Max output
128,000 tokens
Providers
1

Mai Code 1.1 Flash pricing by provider

1 listing across 1 provider. Token prices are in US dollars per 1M tokens. The highlighted row is the price quoted above.

ProviderModel name in LiteLLMContext
GitHub Copilotgithub_copilot/mai-code-1.1-flash128K

Mai Code 1.1 Flash features

Use Mai Code 1.1 Flash with LiteLLM

Call Mai Code 1.1 Flash through the LiteLLM Python SDK, or put it behind the LiteLLM proxy and send OpenAI-format requests to /v1/responses. LiteLLM tracks the cost of every request at these prices.

Python SDK
from litellm import completion

response = completion(
    model="github_copilot/mai-code-1.1-flash",
    messages=[{"role": "user", "content": "Hello!"}],
)
Call the proxy
curl http://0.0.0.0:4000/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer sk-1234" \
  -d '{
    "model": "mai-code-1.1-flash",
    "messages": [{"role": "user", "content": "Hello!"}]
  }'

Mai Code 1.1 Flash FAQ

What is the context window of Mai Code 1.1 Flash?

Mai Code 1.1 Flash has a 128,000-token context window and returns up to 128,000 tokens per response.

What features does Mai Code 1.1 Flash support?

Mai Code 1.1 Flash supports function calling, parallel tool calls, structured output and image input.

Sources

Prices come from LiteLLM's open-source model_prices_and_context_window.json, which LiteLLM uses to calculate the cost of every request.

LiteLLM docs: GitHub Copilot.

Provider pages: GitHub Copilot.