Skip to content
Dashboard

Gemma 4 31B IT

Gemma 4 31B IT is Google's open-weight dense model with 31B parameters, all active during inference. Built on the Gemini 3 architecture, it targets higher output quality than its MoE sibling, with support for function-calling, structured JSON output, native vision, and 140+ languages. Your use is subject to Google's Terms & Privacy Policies.

File InputReasoningTool UseVision (Image)
import { streamText } from 'ai'
const result = streamText({
model: 'google/gemma-4-31b-it',
prompt: 'Why is the sky blue?'
})
Read docs

Copy link to headingProviders

Route requests across multiple providers. Copy a provider slug to set your preference. Visit the docs for more info. Using a provider means you agree to their terms, listed under Legal.

Provider
Context
Max Output
Latency
Throughput
Input
Output
Cache
Web Search
Capabilities
ZDR
No Training
Release Date
262K131K
0.7s
46tps
$0.14/M
$0.40/M
+1
04/02/2026
262K131K
2.4s
16tps
$0.14/M
$0.40/M
04/02/2026
256K131K
2.6s
8tps
$0.39/M
$0.97/M
+1
04/02/2026
131K40K
1.7s
446tps
$0.99/M
$1.49/M
+1
04/02/2026
256K131K
5.2s
81tps
$0.14/M
$0.40/M
+1
04/02/2026
1M1M
1.6s
45tps
$0.14/M
$0.40/M
Read:$0.1/M
Write:
+1
04/02/2026