# Gemma 4 26B A4B IT API Pricing

As of October 10, 2026, Gemma 4 26B A4B IT is free to use on Google AI Studio. Across 8 providers on LiteLLM, the input price ranges from $0 to $0.25. It has a 262,144-token context window and returns up to 32,768 tokens per response.

- Page: https://models.litellm.ai/models/gemma-4-26b-a4b-it
- Lab: Google
- LiteLLM model name: `gemini/gemma-4-26b-a4b-it`
- Context window: 262,144 tokens
- Max output: 32,768 tokens
- Added to LiteLLM: June 2026
- Prices updated: October 10, 2026

## Pricing by provider

Token prices in USD per 1M tokens.

| Provider | Model name | Input | Output | Cache read | Per image | Per second | Context |
|---|---|---|---|---|---|---|---|
| Google AI Studio | `gemini/gemma-4-26b-a4b-it` | $0 | $0 | – | – | – | 256K |
| AIHubMix | `aihubmix/gemma-4-26b-a4b-it` | $0.14 | $0.40 | – | – | – | 256K |
| Cloudflare Workers AI | `cloudflare/@cf/google/gemma-4-26b-a4b-it` | $0.10 | $0.30 | – | – | – | 256K |
| DeepInfra | `deepinfra/google/gemma-4-26B-A4B-it` | $0.07 | $0.34 | – | – | – | 256K |
| Google Vertex AI | `vertex_ai/gemma-4-26b-a4b-it` | $0.15 | $0.60 | $0.015 | – | – | – |
| Google Vertex AI | `vertex_ai/google/gemma-4-26b-a4b-it-maas` | $0.15 | $0.60 | $0.015 | – | – | 256K |
| Novita AI | `novita/google/gemma-4-26b-a4b-it` | $0.13 | $0.40 | – | – | – | 256K |
| Scaleway | `scaleway/google/gemma-4-26b-a4b-it` | $0.25 | $0.50 | – | – | – | 256K |
| Weights & Biases | `wandb/google/gemma-4-26B-A4B-it` | $0.10 | $0.30 | $0.05 | – | – | 262K |

## Cost examples

1,000 requests with 2,000 input and 500 output tokens each:

- Google AI Studio: $0
- DeepInfra: $0.31
- Cloudflare Workers AI: $0.35
- Weights & Biases: $0.35

100 long-document requests with 100,000 input and 2,000 output tokens each:

- Google AI Studio: $0
- DeepInfra: $0.768
- Cloudflare Workers AI: $1.06
- Weights & Biases: $1.06

## Features

Supported: Function calling, Structured output, Image input, Reasoning, Prompt caching.

## Use it with LiteLLM

```python
from litellm import completion

response = completion(
    model="gemini/gemma-4-26b-a4b-it",
    messages=[{"role": "user", "content": "Hello!"}],
)
```

## FAQ

### How much does Gemma 4 26B A4B IT cost?

Gemma 4 26B A4B IT is free to use on Google AI Studio. Across 8 providers on LiteLLM, the input price ranges from $0 to $0.25.

### What is the context window of Gemma 4 26B A4B IT?

Gemma 4 26B A4B IT has a 262,144-token context window and returns up to 32,768 tokens per response.

### Which providers offer Gemma 4 26B A4B IT?

Gemma 4 26B A4B IT is available from 8 providers through LiteLLM: Google AI Studio, AIHubMix, Cloudflare Workers AI, DeepInfra, Google Vertex AI, Novita AI, Scaleway and Weights & Biases. The lowest price is on Google AI Studio.

### How do I call Gemma 4 26B A4B IT with an OpenAI-compatible API?

Use the LiteLLM Python SDK with model="gemini/gemma-4-26b-a4b-it", or add the model to the LiteLLM proxy's config.yaml and send requests to /v1/chat/completions from any OpenAI client. LiteLLM tracks the cost of every request at the prices on this page.

### What features does Gemma 4 26B A4B IT support?

Gemma 4 26B A4B IT supports function calling, structured output, image input, reasoning and prompt caching.

### When did LiteLLM add Gemma 4 26B A4B IT?

Gemma 4 26B A4B IT was added to LiteLLM's model price file in June 2026.

Source: LiteLLM model price file, https://github.com/BerriAI/litellm/blob/main/model_prices_and_context_window.json
