# Gemini 2.5 Flash TTS

> Gemini 2.5 Flash TTS is a proprietary speech model from Google DeepMind, released on 30 Sept 2025. It accepts text and generates audio, with a 33K-token context window and up to 16K output tokens per response. As of 4 Oct 2026, the lowest listed API price is $0.50 per million input tokens and $10 per million output tokens (Google Vertex AI), across 1 provider we track.

Source page: https://www.aimodel.directory/models/gemini-2-5-flash-tts
Last verified: 4 Oct 2026

## Specifications

- **Developer:** Google DeepMind
- **Release date:** 30 Sept 2025 (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google-vertex/models/gemini-2.5-flash-tts.toml)
- **Status:** Generally available (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google-vertex/models/gemini-2.5-flash-tts.toml)
- **Weights:** Proprietary (API only) (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google-vertex/models/gemini-2.5-flash-tts.toml)
- **Context window:** 33K (32,768 tokens) (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google-vertex/models/gemini-2.5-flash-tts.toml)
- **Max output:** 16K tokens (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google-vertex/models/gemini-2.5-flash-tts.toml)
- **Knowledge cutoff:** Jan 2025 (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google-vertex/models/gemini-2.5-flash-tts.toml)
- **Input:** Text (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google-vertex/models/gemini-2.5-flash-tts.toml)
- **Output:** Audio (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google-vertex/models/gemini-2.5-flash-tts.toml)
- **Reasoning mode:** No (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google-vertex/models/gemini-2.5-flash-tts.toml)
- **Tool calling:** No (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google-vertex/models/gemini-2.5-flash-tts.toml)

## API pricing (USD per 1M tokens)

| Provider | Input | Output | Cached input | Context |
|---|---|---|---|---|
| Google Vertex AI | $0.50 | $10 | – | 33K |

## FAQ

### What is the context window of Gemini 2.5 Flash TTS?

Gemini 2.5 Flash TTS has a context window of 32,768 tokens (33K) and can generate up to 16,384 tokens in a single response, according to models.dev.

### How much does the Gemini 2.5 Flash TTS API cost?

Gemini 2.5 Flash TTS is listed from $0.50 per million input tokens and $10 per million output tokens (Google Vertex AI). Prices last verified 4 Oct 2026.

### Is Gemini 2.5 Flash TTS open source?

No. Gemini 2.5 Flash TTS is a proprietary model; Google DeepMind has not released its weights. It is available through APIs including Google Vertex AI.

### When was Gemini 2.5 Flash TTS released?

Google DeepMind released Gemini 2.5 Flash TTS on 30 Sept 2025, according to models.dev.

---
Data licensed CC BY 4.0. Cite as: AI Model Directory, "Gemini 2.5 Flash TTS", https://www.aimodel.directory/models/gemini-2-5-flash-tts