Gemini 3.5 Transcribe API Pricing
As of October 10, 2026, Gemini 3.5 Transcribe costs $2.00 per 1M tokens on Google AI Studio. It has a 98,304-token context window and returns up to 32,768 tokens per response.
- Input
- $2.00 / 1M tokens
- Output
- $12.00 / 1M tokens
- Context window
- 98,304 tokens
- Max output
- 32,768 tokens
- Providers
- 2
- Added to LiteLLM
- August 27, 2026
Gemini 3.5 Transcribe pricing by provider
2 listings across 2 providers. Token prices are in US dollars per 1M tokens. The highlighted row is the price quoted above.
| Provider | Model name in LiteLLM | Per minute | Input | Output | Context |
|---|---|---|---|---|---|
| Google AI Studio | gemini/gemini-3.5-transcribe | – | $2.00 | $12.00 | 96K |
| Google Vertex AI | vertex_ai/gemini-3.5-transcribe | $0.003 | – | $12.00 | – |
Other Gemini 3.5 Transcribe prices on Google AI Studio
| Audio input | $2.00 per 1M tokens |
Gemini 3.5 Transcribe features
Input: text, audio. Output: text.
- Function calling
- Audio input
Use Gemini 3.5 Transcribe with LiteLLM
Call Gemini 3.5 Transcribe through the LiteLLM Python SDK, or put it behind the LiteLLM proxy and send OpenAI-format requests to /v1/audio/transcriptions. LiteLLM tracks the cost of every request at these prices.
from litellm import transcription
response = transcription(
model="gemini/gemini-3.5-transcribe",
file=open("audio.mp3", "rb"),
)model_list:
- model_name: gemini-3.5-transcribe
litellm_params:
model: gemini/gemini-3.5-transcribe
api_key: os.environ/GEMINI_API_KEYcurl http://0.0.0.0:4000/v1/audio/transcriptions \ -H "Authorization: Bearer sk-1234" \ -F model="gemini-3.5-transcribe" \ -F file=@audio.mp3
Gemini 3.5 Transcribe FAQ
How much does Gemini 3.5 Transcribe cost?
Gemini 3.5 Transcribe costs $2.00 per 1M tokens on Google AI Studio.
What is the context window of Gemini 3.5 Transcribe?
Gemini 3.5 Transcribe has a 98,304-token context window and returns up to 32,768 tokens per response.
Which providers offer Gemini 3.5 Transcribe?
Gemini 3.5 Transcribe is available from 2 providers through LiteLLM: Google AI Studio and Google Vertex AI.
How do I call Gemini 3.5 Transcribe with an OpenAI-compatible API?
Use the LiteLLM Python SDK with model="gemini/gemini-3.5-transcribe", or add the model to the LiteLLM proxy's config.yaml and send requests to /v1/audio/transcriptions from any OpenAI client. LiteLLM tracks the cost of every request at the prices on this page.
What features does Gemini 3.5 Transcribe support?
Gemini 3.5 Transcribe supports audio input. It does not support function calling.
When did LiteLLM add Gemini 3.5 Transcribe?
Gemini 3.5 Transcribe was added to LiteLLM's model price file on August 27, 2026.
Sources
Prices come from LiteLLM's open-source model_prices_and_context_window.json, which LiteLLM uses to calculate the cost of every request. Provider pricing pages:
- ai.google.dev/gemini-api/docs/pricing
- cloud.google.com/gemini-enterprise-agent-platform/generative-ai/pricing
LiteLLM docs: Google AI Studio, Google Vertex AI.
Provider pages: Google AI Studio, Google Vertex AI.
