# Gemini 3.7 Flash

> Gemini 3.7 Flash is a proprietary language model from Google DeepMind, released on 13 Aug 2026. It accepts text, images, video, audio and PDFs and generates text, with a 1M-token context window and up to 66K output tokens per response. As of 4 Oct 2026, the lowest listed API price is $0.75 per million input tokens and $3.75 per million output tokens (Gemini API), across 4 providers we track.

Source page: https://www.aimodel.directory/models/gemini-3-7-flash
Last verified: 4 Oct 2026

## Specifications

- **Developer:** Google DeepMind
- **Release date:** 13 Aug 2026 (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.7-flash.toml)
- **Status:** Generally available (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.7-flash.toml)
- **Weights:** Proprietary (API only) (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.7-flash.toml)
- **Context window:** 1M (1,048,576 tokens) (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.7-flash.toml)
- **Max output:** 66K tokens (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.7-flash.toml)
- **Knowledge cutoff:** Mar 2026 (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.7-flash.toml)
- **Input:** Text, images, video, audio and PDFs (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.7-flash.toml)
- **Output:** Text (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.7-flash.toml)
- **Reasoning mode:** Yes (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.7-flash.toml)
- **Tool calling:** Yes (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.7-flash.toml)
- **Structured output:** Yes (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.7-flash.toml)

## API pricing (USD per 1M tokens)

| Provider | Input | Output | Cached input | Context |
|---|---|---|---|---|
| Gemini API | $0.75 | $3.75 | $0.07 | 1M |
| Google Vertex AI | $0.75 | $3.75 | $0.07 | 1M |
| OpenRouter | $0.75 | $3.75 | $0.07 | 1M |
| Vercel AI Gateway | $0.75 | $3.75 | $0.07 | 1M |

## FAQ

### What is the context window of Gemini 3.7 Flash?

Gemini 3.7 Flash has a context window of 1,048,576 tokens (1M) and can generate up to 65,536 tokens in a single response, according to models.dev.

### How much does the Gemini 3.7 Flash API cost?

Through Gemini API, Gemini 3.7 Flash costs $0.75 per million input tokens and $3.75 per million output tokens. Prices last verified 4 Oct 2026.

### Is Gemini 3.7 Flash open source?

No. Gemini 3.7 Flash is a proprietary model; Google DeepMind has not released its weights. It is available through APIs including Gemini API, Google Vertex AI, OpenRouter, Vercel AI Gateway.

### When was Gemini 3.7 Flash released?

Google DeepMind released Gemini 3.7 Flash on 13 Aug 2026, according to models.dev.

### Which providers offer Gemini 3.7 Flash?

We track Gemini 3.7 Flash on 4 providers: Gemini API, Google Vertex AI, OpenRouter, Vercel AI Gateway.

---
Data licensed CC BY 4.0. Cite as: AI Model Directory, "Gemini 3.7 Flash", https://www.aimodel.directory/models/gemini-3-7-flash