# Gemini 3.1 Flash Live Preview

> Gemini 3.1 Flash Live Preview is a proprietary language model from Google DeepMind, released on 26 Mar 2026. It accepts text, images, video and audio and generates text and audio, with a 131K-token context window and up to 66K output tokens per response. As of 4 Oct 2026, the lowest listed API price is $0.75 per million input tokens and $4.50 per million output tokens (Gemini API), across 1 provider we track.

Source page: https://www.aimodel.directory/models/gemini-3-1-flash-live-preview
Last verified: 4 Oct 2026

## Specifications

- **Developer:** Google DeepMind
- **Release date:** 26 Mar 2026 (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.1-flash-live-preview.toml)
- **Status:** Generally available (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.1-flash-live-preview.toml)
- **Weights:** Proprietary (API only) (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.1-flash-live-preview.toml)
- **Context window:** 131K (131,072 tokens) (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.1-flash-live-preview.toml)
- **Max output:** 66K tokens (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.1-flash-live-preview.toml)
- **Knowledge cutoff:** Jan 2025 (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.1-flash-live-preview.toml)
- **Input:** Text, images, video and audio (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.1-flash-live-preview.toml)
- **Output:** Text and audio (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.1-flash-live-preview.toml)
- **Reasoning mode:** Yes (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.1-flash-live-preview.toml)
- **Tool calling:** Yes (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.1-flash-live-preview.toml)
- **Structured output:** No (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.1-flash-live-preview.toml)

## API pricing (USD per 1M tokens)

| Provider | Input | Output | Cached input | Context |
|---|---|---|---|---|
| Gemini API | $0.75 | $4.50 | – | 131K |

## FAQ

### What is the context window of Gemini 3.1 Flash Live Preview?

Gemini 3.1 Flash Live Preview has a context window of 131,072 tokens (131K) and can generate up to 65,536 tokens in a single response, according to models.dev.

### How much does the Gemini 3.1 Flash Live Preview API cost?

Through Gemini API, Gemini 3.1 Flash Live Preview costs $0.75 per million input tokens and $4.50 per million output tokens. Prices last verified 4 Oct 2026.

### Is Gemini 3.1 Flash Live Preview open source?

No. Gemini 3.1 Flash Live Preview is a proprietary model; Google DeepMind has not released its weights. It is available through APIs including Gemini API.

### When was Gemini 3.1 Flash Live Preview released?

Google DeepMind released Gemini 3.1 Flash Live Preview on 26 Mar 2026, according to models.dev.

---
Data licensed CC BY 4.0. Cite as: AI Model Directory, "Gemini 3.1 Flash Live Preview", https://www.aimodel.directory/models/gemini-3-1-flash-live-preview