[Google](/ai-gateway/models/labs/google)

# Gemma 4 26B A4B IT

Gemma 4 26B A4B IT is Google's open-weight mixture-of-experts model with 26B total parameters and roughly 4B active per forward pass. Built on the Gemini 3 architecture, it supports function-calling, structured JSON output, native vision, and 140+ languages within a context window of 1.0M tokens. Your use is subject to Google's [Terms](https://policies.google.com/terms/generative-ai) & [Privacy](https://policies.google.com/privacy) Policies.

File InputReasoningTool UseVision (Image)

[Use with AI Gateway](https://vercel.com/d?to=%2F%5Bteam%5D%2F%7E%2Fai%3Futm_source%3Dgateway-model-page%26utm_campaign%3Dai-gateway-models&title=Get+Started+with+Vercel+AI+Gateway) [View docs](https://vercel.com/docs/ai-gateway)

AI SDKChat CompletionsMessagesResponses

```
1import { streamText } from 'ai'
2

3const result = streamText({
4  model: 'google/gemma-4-26b-a4b-it',
5  prompt: 'Why is the sky blue?'
6})
```

[Read docs](https://vercel.com/docs/ai-gateway/sdks-and-apis/ai-sdk)

[Overview](/ai-gateway/models/gemma-4-26b-a4b-it) [API](/ai-gateway/models/gemma-4-26b-a4b-it/api) [About](/ai-gateway/models/gemma-4-26b-a4b-it/about) [Providers](/ai-gateway/models/gemma-4-26b-a4b-it/providers) [Throughput](/ai-gateway/models/gemma-4-26b-a4b-it/throughput) [Latency](/ai-gateway/models/gemma-4-26b-a4b-it/latency) [Uptime](/ai-gateway/models/gemma-4-26b-a4b-it/uptime) [Status](/ai-gateway/models/gemma-4-26b-a4b-it/status) [Similar](/ai-gateway/models/gemma-4-26b-a4b-it/similar) [FAQ](/ai-gateway/models/gemma-4-26b-a4b-it/faq)

## [Copy link to heading](#playground)Playground

Try out Gemma 4 26B A4B IT by Google. Usage is billed to your team at API rates. Free users (those who haven't made a payment) get $5 of credits every 30 days.

![google logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fgoogle.png&w=48&q=75)Gemma 4 26B A4B IT

![google logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fgoogle.png&w=96&q=75)

Gemma 4 26B A4B IT

## [Copy link to heading](#providers)Providers

Route requests across multiple providers. Copy a provider slug to set your preference. Visit the [docs](/docs/ai-gateway/provider-options) for more info. Using a provider means you agree to their terms, listed under Legal.

| Provider |
| --- |

| Context | Max Output | Latency | Throughput | Input | Output | Cache | Web Search | Capabilities | ZDR | No Training | Release Date |  |
| --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- |

| ![novita logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fnovita.png&w=48&q=75) [Novita AI](/ai-gateway/models/providers/novita) Legal:[Terms](https://novita.ai/legal/terms-of-service)•[Privacy](https://novita.ai/legal/privacy-policy) | 262K | 131K | 0.6s | 29tps | $0.13/M | $0.40/M |  | — | +1 |  |  | 04/02/2026 |  |
| --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- |
| ![parasail logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fparasail.png&w=48&q=75) [Parasail](/ai-gateway/models/providers/parasail) Legal:[Terms](https://www.parasail.io/legal/terms)•[Privacy](https://www.saas.parasail.io/assets/privacy-policy.html) | 262K | 131K | 0.6s | 90tps | $0.13/M | $0.40/M |  | — |  |  |  | 04/02/2026 |  |
| ![vertex logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fvertex%2520ai.png&w=48&q=75) [Google Vertex AI](/ai-gateway/models/providers/vertex) Legal:[Terms](https://cloud.google.com/terms/service-terms)•[Privacy](https://cloud.google.com/privacy) | 262K | 131K | 0.3s | 87tps | $0.15/M | $0.60/M | Read:$0.01/M Write:— | — | +2 |  |  | 04/02/2026 |  |
| ![gmicloud logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fgmicloud.png%3Fv%3D1785888489595&w=48&q=75) [GMICloud](/ai-gateway/models/providers/gmicloud) Legal:[Terms](https://www.gmicloud.ai/en/terms-and-conditions)•[Privacy](https://www.gmicloud.ai/en/privacy-policy) | 1M | 1M | 0.8s | 49tps | $0.13/M | $0.40/M |  | — | +1 |  |  | 04/02/2026 |  |

## [Copy link to heading](#throughput)Throughput24 hours

1W

1D

P50 throughput on live AI Gateway traffic, in tokens per second (TPS). Visit the [docs](https://vercel.com/docs/ai-gateway/models-and-providers/metrics) for more info.

## [Copy link to heading](#latency)Latency24 hours

1W

1D

P50 time to first token (TTFT) on live AI Gateway traffic, in milliseconds. View the [docs](https://vercel.com/docs/ai-gateway/models-and-providers/metrics) for more info.

## [Copy link to heading](#uptime)Uptime24 hours

1W

1D

1H

Direct request success rate on AI Gateway and per-provider. Visit the [docs](https://vercel.com/docs/ai-gateway/models-and-providers/uptime) for more info.

1W

1D

1H

## [Copy link to heading](#more-models-by-google)More models by Google

All

Text

Code

Video Input

| Model |
| --- |

| Context | Latency | Throughput | Input | Output | Cache | Web Search | Capabilities | Providers | ZDR | No Training | Release Date |  |
| --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- |

| ![google logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fgoogle.png&w=48&q=75) [google/gemini-3.7-flash](/ai-gateway/models/gemini-3.7-flash) | 1M | 1.6s | 565tps | $1.50/M$0.75/M | $7.50/M$3.75/M | Read:$0.15/M$0.07/M Write:— | — | +3 | ![google logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fgoogle.png&w=48&q=75) ![vertex logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fvertex%2520ai.png&w=48&q=75) |  |  | 08/13/2026 |  |
| --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- |
| ![google logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fgoogle.png&w=48&q=75) [google/gemini-3.5-flash-lite](/ai-gateway/models/gemini-3.5-flash-lite) | 1M | 0.6s | 390tps | $0.30/M | $2.50/M | Read:$0.03/M Write:— | $14/K+1 more \+ input costs | +3 | ![google logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fgoogle.png&w=48&q=75) ![vertex logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fvertex%2520ai.png&w=48&q=75) |  |  | 07/21/2026 |  |
| ![google logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fgoogle.png&w=48&q=75) [google/gemini-3.5-flash](/ai-gateway/models/gemini-3.5-flash) | 1M | 2.8s | 192tps | $1.50/M | $9/M | Read:$0.15/M Write:— | $14/K+1 more \+ input costs | +3 | ![google logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fgoogle.png&w=48&q=75) ![vertex logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fvertex%2520ai.png&w=48&q=75) |  |  | 05/19/2026 |  |
| ![google logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fgoogle.png&w=48&q=75) [google/gemini-3.1-flash-lite](/ai-gateway/models/gemini-3.1-flash-lite) | 1M | 0.5s | 372tps | $0.25/M | $1.50/M | Read:$0.03/M Write:— | $14/K+1 more \+ input costs | +3 | ![google logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fgoogle.png&w=48&q=75) ![vertex logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fvertex%2520ai.png&w=48&q=75) |  |  | 05/07/2026 |  |
| ![google logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fgoogle.png&w=48&q=75) [google/gemini\-3-flash](/ai-gateway/models/gemini-3-flash) | 1M | 0.6s | 192tps | $0.50/M+1 more | $3/M+1 more | Read: $0.05/M+1 more Write: — | $14/K+1 more \+ input costs | +3 | ![google logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fgoogle.png&w=48&q=75) ![vertex logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fvertex%2520ai.png&w=48&q=75) |  |  | 12/17/2025 |  |
| ![google logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fgoogle.png&w=48&q=75) [google/gemini-2.5-flash-lite](/ai-gateway/models/gemini-2.5-flash-lite) | 1M | 0.2s | 414tps | $0.10/M | $0.40/M | Read:$0.01/M Write:— | $35/K+1 more \+ input costs | +3 | ![google logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fgoogle.png&w=48&q=75) ![vertex logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fvertex%2520ai.png&w=48&q=75) |  |  | 06/17/2025 |  |

## [Copy link to heading](#about-gemma-4-26b-a4b-it)About Gemma 4 26B A4B IT

Gemma 4 26B A4B IT is part of Google's Gemma 4 family, the open-weight counterpart to the proprietary Gemini lineup. Google released it on April 2, 2026 as an instruction-tuned mixture-of-experts (MoE) model built on the same architecture as Gemini 3.

The MoE design is the defining characteristic. Of the 26B total parameters, only roughly 4B are active during any single forward pass. A routing mechanism selects which expert sub-networks to activate for each input, so Gemma 4 26B A4B IT achieves quality comparable to a much larger dense model while using a fraction of the compute per token. This translates to lower latency and higher throughput. See live metrics on this page for current throughput.

Gemma 4 26B A4B IT accepts text and image inputs within a context window of 1.0M tokens and supports over 140 languages. It handles function-calling, agentic workflows, structured JSON output, and system instructions natively. The instruction-tuning (indicated by the `it` suffix) means Gemma 4 26B A4B IT is ready for conversational and task-oriented use out of the box.

Running Gemma 4 26B A4B IT through AI Gateway provides unified billing, observability, automatic retries, and provider failover without requiring infrastructure management.

## [Copy link to heading](#what-to-consider-when-choosing-a-provider)What To Consider When Choosing a Provider

- Configuration: Evaluate whether the MoE architecture's latency and throughput characteristics fit your workload before selecting a provider variant at production scale.
- Zero Data Retention: Zero Data Retention is available for this model. It is offered on a per-provider and model basis. See the [documentation](https://vercel.com/docs/ai-gateway/security-and-compliance/zdr) for details.
- Authentication: AI Gateway authenticates requests using an [API key](https://vercel.com/docs/ai-gateway/authentication-and-byok#api-key-authentication) or [OIDC token](https://vercel.com/docs/ai-gateway/authentication-and-byok#oidc-token-authentication). You do not need to manage provider credentials directly.

## [Copy link to heading](#when-to-use-gemma-4-26b-a4b-it)When to Use Gemma 4 26B A4B IT

### Best for

- Latency-sensitive production workloads: The MoE architecture's lower compute-per-token translates to faster response times
- Cost-efficient agentic pipelines: Need function-calling and structured output at high request volumes
- Multilingual applications: Serving users across 140+ languages with a single model
- Vision-language tasks: Image understanding, visual Q&A, and document analysis within a context window of 1.0M tokens
- Open-weight workloads: The ability to inspect model weights matters

### Consider alternatives when

- Highest output quality needed: Latency is not a constraint, and the dense Gemma 4 31B or a Gemini model may be more appropriate
- Native image or audio generation: Your task requires media output, which Gemma 4 26B A4B IT does not support
- Simple classification or extraction: A smaller, cheaper model is sufficient for straightforward workloads

## [Copy link to heading](#conclusion)Conclusion

Gemma 4 26B A4B IT provides Gemini 3-class capabilities in an open-weight package optimized for throughput. The MoE architecture keeps inference fast and affordable. For teams that need strong multilingual, multimodal reasoning without proprietary lock-in, it is a practical production choice on AI Gateway.