Google DeepMind

Gemma 4 31B IT

Generally availableOpen weights262K contextReasoning

Summary

Gemma 4 31B IT is an open-weight language model from Google DeepMind, released on 2 Apr 2026. It accepts text and images and generates text, with a 262K-token context window and up to 33K output tokens per response. As of 4 Oct 2026, the lowest listed API price is $0.09 per million input tokens and $0.34 per million output tokens (OpenRouter), across 10 providers we track.

Largest Gemma 4 instruction model for open, self-hosted chat and reasoning

Last verified 4 Oct 2026. Every figure on this page links to its source.

Specifications

DeveloperGoogle DeepMind
Release date2 Apr 2026 [models.dev]
StatusGenerally available [models.dev]
WeightsOpen weights [Hugging Face]
Licenseapache-2.0 [Hugging Face]
Parameters31.3B [Hugging Face]
Context window262K (262,144 tokens) [models.dev]
Max output33K tokens [models.dev]
Knowledge cutoffJan 2025 [models.dev]
InputText and images [models.dev]
OutputText [models.dev]
Reasoning modeYes [models.dev]
Tool callingYes [models.dev]
Structured outputYes [models.dev]
Hugging Facegoogle/gemma-4-31B-it [Hugging Face]

Gemma 4 31B IT API pricing by provider

USD per million tokens. Sorted by input price.

ProviderInputOutputCached inputContextSource
NVIDIA NIM
google/gemma-4-31b-it
FreeFree–256Kmodels.dev, 4 Oct 2026
OpenRouter
google/gemma-4-31b-it
$0.09$0.34$0.05262KOpenRouter, 4 Oct 2026
SiliconFlow
google/gemma-4-31B-it
$0.13$0.40–262Kmodels.dev, 4 Oct 2026
Amazon Bedrock
google.gemma-4-31b
$0.14$0.40–262Kmodels.dev, 4 Oct 2026
FriendliAI
google/gemma-4-31B-it
$0.14$0.40–262Kmodels.dev, 4 Oct 2026
Hugging Face Inference Providers
google/gemma-4-31B-it
$0.14$0.40–262Kmodels.dev, 4 Oct 2026
Novita AI
google/gemma-4-31b-it
$0.14$0.40–262Kmodels.dev, 4 Oct 2026
Vercel AI Gateway
google/gemma-4-31b-it
$0.14$0.40–262Kmodels.dev, 4 Oct 2026
DeepInfra
google/gemma-4-31B-it
$0.20$0.40–262Kmodels.dev, 4 Oct 2026
Gemini API
gemma-4-31b-it
–––262Kmodels.dev, 4 Oct 2026

Frequently asked questions

What is the context window of Gemma 4 31B IT?

Gemma 4 31B IT has a context window of 262,144 tokens (262K) and can generate up to 32,768 tokens in a single response, according to models.dev.

How much does the Gemma 4 31B IT API cost?

Gemma 4 31B IT is listed from $0.09 per million input tokens and $0.34 per million output tokens (OpenRouter). Prices last verified 4 Oct 2026.

Is Gemma 4 31B IT open source?

Gemma 4 31B IT is an open-weight model: its weights are published on Hugging Face as google/gemma-4-31B-it under the apache-2.0 license. Check the license terms before commercial use.

When was Gemma 4 31B IT released?

Google DeepMind released Gemma 4 31B IT on 2 Apr 2026, according to models.dev.

Which providers offer Gemma 4 31B IT?

We track Gemma 4 31B IT on 10 providers: NVIDIA NIM, OpenRouter, SiliconFlow, Amazon Bedrock, FriendliAI, Hugging Face Inference Providers, Novita AI, Vercel AI Gateway, DeepInfra, Gemini API.

Change history

  • DeepInfra price changed: input $0.15 → $0.20, output $0.40 → $0.40 per 1M tokens

Sources