# Gemini 3.1 Flash Lite Preview

> Gemini 3.1 Flash Lite Preview is a proprietary language model from Google DeepMind, released on 3 Mar 2026. It accepts text, images, video, audio and PDFs and generates text, with a 1M-token context window and up to 66K output tokens per response. As of 4 Oct 2026, the lowest listed API price is $0.25 per million input tokens and $1.50 per million output tokens (Databricks), across 2 providers we track. Google DeepMind has deprecated this model.

Source page: https://www.aimodel.directory/models/gemini-3-1-flash-lite-preview
Last verified: 4 Oct 2026

## Specifications

- **Developer:** Google DeepMind
- **Release date:** 3 Mar 2026 (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.1-flash-lite-preview.toml)
- **Status:** Deprecated (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.1-flash-lite-preview.toml)
- **Weights:** Proprietary (API only) (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.1-flash-lite-preview.toml)
- **Context window:** 1M (1,048,576 tokens) (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.1-flash-lite-preview.toml)
- **Max output:** 66K tokens (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.1-flash-lite-preview.toml)
- **Knowledge cutoff:** Jan 2025 (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.1-flash-lite-preview.toml)
- **Input:** Text, images, video, audio and PDFs (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.1-flash-lite-preview.toml)
- **Output:** Text (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.1-flash-lite-preview.toml)
- **Reasoning mode:** Yes (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.1-flash-lite-preview.toml)
- **Tool calling:** Yes (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.1-flash-lite-preview.toml)
- **Structured output:** Yes (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-3.1-flash-lite-preview.toml)

## API pricing (USD per 1M tokens)

| Provider | Input | Output | Cached input | Context |
|---|---|---|---|---|
| Databricks | $0.25 | $1.50 | $0.03 | 1M |
| OpenRouter | $0.25 | $1.50 | $0.03 | 1M |

## FAQ

### What is the context window of Gemini 3.1 Flash Lite Preview?

Gemini 3.1 Flash Lite Preview has a context window of 1,048,576 tokens (1M) and can generate up to 65,536 tokens in a single response, according to models.dev.

### How much does the Gemini 3.1 Flash Lite Preview API cost?

Gemini 3.1 Flash Lite Preview is listed from $0.25 per million input tokens and $1.50 per million output tokens (Databricks). Prices last verified 4 Oct 2026.

### Is Gemini 3.1 Flash Lite Preview open source?

No. Gemini 3.1 Flash Lite Preview is a proprietary model; Google DeepMind has not released its weights. It is available through APIs including Databricks, OpenRouter.

### When was Gemini 3.1 Flash Lite Preview released?

Google DeepMind released Gemini 3.1 Flash Lite Preview on 3 Mar 2026, according to models.dev.

### Which providers offer Gemini 3.1 Flash Lite Preview?

We track Gemini 3.1 Flash Lite Preview on 2 providers: Databricks, OpenRouter.

---
Data licensed CC BY 4.0. Cite as: AI Model Directory, "Gemini 3.1 Flash Lite Preview", https://www.aimodel.directory/models/gemini-3-1-flash-lite-preview