[Google](/ai-gateway/models/labs/google)

# Gemma 4 31B IT

Gemma 4 31B IT is Google's open-weight dense model with 31B parameters, all active during inference. Built on the Gemini 3 architecture, it targets higher output quality than its MoE sibling, with support for function-calling, structured JSON output, native vision, and 140+ languages. Your use is subject to Google's [Terms](https://policies.google.com/terms/generative-ai) & [Privacy](https://policies.google.com/privacy) Policies.

File InputReasoningTool UseVision (Image)

[Use with AI Gateway](https://vercel.com/d?to=%2F%5Bteam%5D%2F%7E%2Fai%3Futm_source%3Dgateway-model-page%26utm_campaign%3Dai-gateway-models&title=Get+Started+with+Vercel+AI+Gateway) [View docs](https://vercel.com/docs/ai-gateway)

AI SDKChat CompletionsMessagesResponses

```
1import { streamText } from 'ai'
2

3const result = streamText({
4  model: 'google/gemma-4-31b-it',
5  prompt: 'Why is the sky blue?'
6})
```

[Read docs](https://vercel.com/docs/ai-gateway/sdks-and-apis/ai-sdk)

[Overview](/ai-gateway/models/gemma-4-31b-it) [API](/ai-gateway/models/gemma-4-31b-it/api) [About](/ai-gateway/models/gemma-4-31b-it/about) [Providers](/ai-gateway/models/gemma-4-31b-it/providers) [Throughput](/ai-gateway/models/gemma-4-31b-it/throughput) [Latency](/ai-gateway/models/gemma-4-31b-it/latency) [Uptime](/ai-gateway/models/gemma-4-31b-it/uptime) [Status](/ai-gateway/models/gemma-4-31b-it/status) [Similar](/ai-gateway/models/gemma-4-31b-it/similar) [FAQ](/ai-gateway/models/gemma-4-31b-it/faq)

## [Copy link to heading](#latency)Latency24 hours

1W

1D

P50 time to first token (TTFT) on live AI Gateway traffic, in milliseconds. View the [docs](https://vercel.com/docs/ai-gateway/models-and-providers/metrics) for more info.