Gemma 4 31B IT
Summary
Gemma 4 31B IT is an open-weight language model from Google DeepMind, released on 2 Apr 2026. It accepts text and images and generates text, with a 262K-token context window and up to 33K output tokens per response. As of 4 Oct 2026, the lowest listed API price is $0.09 per million input tokens and $0.34 per million output tokens (OpenRouter), across 10 providers we track.
Largest Gemma 4 instruction model for open, self-hosted chat and reasoning
Last verified 4 Oct 2026. Every figure on this page links to its source.
Specifications
| Developer | Google DeepMind |
|---|---|
| Release date | 2 Apr 2026 [models.dev] |
| Status | Generally available [models.dev] |
| Weights | Open weights [Hugging Face] |
| License | apache-2.0 [Hugging Face] |
| Parameters | 31.3B [Hugging Face] |
| Context window | 262K (262,144 tokens) [models.dev] |
| Max output | 33K tokens [models.dev] |
| Knowledge cutoff | Jan 2025 [models.dev] |
| Input | Text and images [models.dev] |
| Output | Text [models.dev] |
| Reasoning mode | Yes [models.dev] |
| Tool calling | Yes [models.dev] |
| Structured output | Yes [models.dev] |
| Hugging Face | google/gemma-4-31B-it [Hugging Face] |
Gemma 4 31B IT API pricing by provider
USD per million tokens. Sorted by input price.
| Provider | Input | Output | Cached input | Context | Source |
|---|---|---|---|---|---|
| NVIDIA NIM google/gemma-4-31b-it | Free | Free | – | 256K | models.dev, 4 Oct 2026 |
| OpenRouter google/gemma-4-31b-it | $0.09 | $0.34 | $0.05 | 262K | OpenRouter, 4 Oct 2026 |
| SiliconFlow google/gemma-4-31B-it | $0.13 | $0.40 | – | 262K | models.dev, 4 Oct 2026 |
| Amazon Bedrock google.gemma-4-31b | $0.14 | $0.40 | – | 262K | models.dev, 4 Oct 2026 |
| FriendliAI google/gemma-4-31B-it | $0.14 | $0.40 | – | 262K | models.dev, 4 Oct 2026 |
| Hugging Face Inference Providers google/gemma-4-31B-it | $0.14 | $0.40 | – | 262K | models.dev, 4 Oct 2026 |
| Novita AI google/gemma-4-31b-it | $0.14 | $0.40 | – | 262K | models.dev, 4 Oct 2026 |
| Vercel AI Gateway google/gemma-4-31b-it | $0.14 | $0.40 | – | 262K | models.dev, 4 Oct 2026 |
| DeepInfra google/gemma-4-31B-it | $0.20 | $0.40 | – | 262K | models.dev, 4 Oct 2026 |
| Gemini API gemma-4-31b-it | – | – | – | 262K | models.dev, 4 Oct 2026 |
Frequently asked questions
What is the context window of Gemma 4 31B IT?
Gemma 4 31B IT has a context window of 262,144 tokens (262K) and can generate up to 32,768 tokens in a single response, according to models.dev.
How much does the Gemma 4 31B IT API cost?
Gemma 4 31B IT is listed from $0.09 per million input tokens and $0.34 per million output tokens (OpenRouter). Prices last verified 4 Oct 2026.
Is Gemma 4 31B IT open source?
Gemma 4 31B IT is an open-weight model: its weights are published on Hugging Face as google/gemma-4-31B-it under the apache-2.0 license. Check the license terms before commercial use.
When was Gemma 4 31B IT released?
Google DeepMind released Gemma 4 31B IT on 2 Apr 2026, according to models.dev.
Which providers offer Gemma 4 31B IT?
We track Gemma 4 31B IT on 10 providers: NVIDIA NIM, OpenRouter, SiliconFlow, Amazon Bedrock, FriendliAI, Hugging Face Inference Providers, Novita AI, Vercel AI Gateway, DeepInfra, Gemini API.
Change history
- DeepInfra price changed: input $0.15 → $0.20, output $0.40 → $0.40 per 1M tokens
Sources
- https://openrouter.ai/google/gemma-4-31b-it
- https://github.com/sst/models.dev/blob/dev/providers/google/models/gemma-4-31b-it.toml
- https://huggingface.co/google/gemma-4-31B-it
- https://github.com/sst/models.dev/blob/dev/providers/nvidia/models/google/gemma-4-31b-it.toml
- https://github.com/sst/models.dev/blob/dev/providers/siliconflow/models/google/gemma-4-31B-it.toml
- https://github.com/sst/models.dev/blob/dev/providers/amazon-bedrock/models/google.gemma-4-31b.toml
- https://github.com/sst/models.dev/blob/dev/providers/friendli/models/google/gemma-4-31B-it.toml
- https://github.com/sst/models.dev/blob/dev/providers/huggingface/models/google/gemma-4-31B-it.toml
- https://github.com/sst/models.dev/blob/dev/providers/novita-ai/models/google/gemma-4-31b-it.toml
- https://github.com/sst/models.dev/blob/dev/providers/vercel/models/google/gemma-4-31b-it.toml
- https://github.com/sst/models.dev/blob/dev/providers/deepinfra/models/google/gemma-4-31B-it.toml