# Gemma 4 12B IT

> Gemma 4 12B IT is a proprietary language model from Google DeepMind, released on 9 Jun 2026. It accepts text and generates text, with a 262K-token context window and up to 262K output tokens per response. As of 4 Oct 2026, the lowest listed API price is $0.10 per million input tokens and $0.30 per million output tokens (SiliconFlow), across 1 provider we track.

Source page: https://www.aimodel.directory/models/gemma-4-12b-it
Last verified: 4 Oct 2026

## Specifications

- **Developer:** Google DeepMind
- **Release date:** 9 Jun 2026 (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/siliconflow/models/google/gemma-4-12B-it.toml)
- **Status:** Generally available (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/siliconflow/models/google/gemma-4-12B-it.toml)
- **Weights:** Proprietary (API only) (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/siliconflow/models/google/gemma-4-12B-it.toml)
- **Context window:** 262K (262,144 tokens) (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/siliconflow/models/google/gemma-4-12B-it.toml)
- **Max output:** 262K tokens (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/siliconflow/models/google/gemma-4-12B-it.toml)
- **Input:** Text (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/siliconflow/models/google/gemma-4-12B-it.toml)
- **Output:** Text (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/siliconflow/models/google/gemma-4-12B-it.toml)
- **Reasoning mode:** No (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/siliconflow/models/google/gemma-4-12B-it.toml)
- **Tool calling:** Yes (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/siliconflow/models/google/gemma-4-12B-it.toml)
- **Structured output:** Yes (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/siliconflow/models/google/gemma-4-12B-it.toml)

## API pricing (USD per 1M tokens)

| Provider | Input | Output | Cached input | Context |
|---|---|---|---|---|
| SiliconFlow | $0.10 | $0.30 | – | 262K |

## FAQ

### What is the context window of Gemma 4 12B IT?

Gemma 4 12B IT has a context window of 262,144 tokens (262K) and can generate up to 262,144 tokens in a single response, according to models.dev.

### How much does the Gemma 4 12B IT API cost?

Gemma 4 12B IT is listed from $0.10 per million input tokens and $0.30 per million output tokens (SiliconFlow). Prices last verified 4 Oct 2026.

### Is Gemma 4 12B IT open source?

No. Gemma 4 12B IT is a proprietary model; Google DeepMind has not released its weights. It is available through APIs including SiliconFlow.

### When was Gemma 4 12B IT released?

Google DeepMind released Gemma 4 12B IT on 9 Jun 2026, according to models.dev.

---
Data licensed CC BY 4.0. Cite as: AI Model Directory, "Gemma 4 12B IT", https://www.aimodel.directory/models/gemma-4-12b-it