Gemma 4 E4B IT
Summary
Gemma 4 E4B IT is an open-weight language model from Google DeepMind, released on 2 Apr 2026. It accepts text, images and audio and generates text, with a 33K-token context window and up to 33K output tokens per response.
Open Gemma instruction model for efficient chat and self-hosted deployments
Last verified 4 Oct 2026. Every figure on this page links to its source.
Specifications
| Developer | Google DeepMind |
|---|---|
| Release date | 2 Apr 2026 [models.dev] |
| Status | Generally available [models.dev] |
| Weights | Open weights [models.dev] |
| Context window | 33K (32,768 tokens) [models.dev] |
| Max output | 33K tokens [models.dev] |
| Input | Text, images and audio [models.dev] |
| Output | Text [models.dev] |
| Reasoning mode | Yes [models.dev] |
| Tool calling | Yes [models.dev] |
| Structured output | Yes [models.dev] |
Gemma 4 E4B IT API pricing by provider
USD per million tokens. Sorted by input price.
We do not currently track a paid API listing for this model.
Frequently asked questions
What is the context window of Gemma 4 E4B IT?
Gemma 4 E4B IT has a context window of 32,768 tokens (33K) and can generate up to 32,768 tokens in a single response, according to models.dev.
Is Gemma 4 E4B IT open source?
Gemma 4 E4B IT is an open-weight model. Check the license terms before commercial use.
When was Gemma 4 E4B IT released?
Google DeepMind released Gemma 4 E4B IT on 2 Apr 2026, according to models.dev.
Change history
No changes detected since we started tracking this model. We re-check every source daily.