# Gemini Embedding 001

> Gemini Embedding 001 is a proprietary language model from Google DeepMind, released on 20 May 2025. It accepts text and generates text, with a 2K-token context window and up to 1 output tokens per response. As of 4 Oct 2026, the lowest listed API price is $0.15 per million input tokens and Free per million output tokens (Gemini API), across 3 providers we track.

Source page: https://www.aimodel.directory/models/gemini-embedding-001
Last verified: 4 Oct 2026

## Specifications

- **Developer:** Google DeepMind
- **Release date:** 20 May 2025 (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-embedding-001.toml)
- **Status:** Generally available (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-embedding-001.toml)
- **Weights:** Proprietary (API only) (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-embedding-001.toml)
- **Context window:** 2K (2,048 tokens) (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-embedding-001.toml)
- **Max output:** 1 tokens (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-embedding-001.toml)
- **Knowledge cutoff:** May 2025 (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-embedding-001.toml)
- **Input:** Text (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-embedding-001.toml)
- **Output:** Text (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-embedding-001.toml)
- **Reasoning mode:** No (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-embedding-001.toml)
- **Tool calling:** No (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/google/models/gemini-embedding-001.toml)

## API pricing (USD per 1M tokens)

| Provider | Input | Output | Cached input | Context |
|---|---|---|---|---|
| Gemini API | $0.15 | Free | – | 2K |
| Google Vertex AI | $0.15 | Free | – | 2K |
| Vercel AI Gateway | – | – | – | 8K |

## FAQ

### What is the context window of Gemini Embedding 001?

Gemini Embedding 001 has a context window of 2,048 tokens (2K) and can generate up to 1 tokens in a single response, according to models.dev.

### How much does the Gemini Embedding 001 API cost?

Through Gemini API, Gemini Embedding 001 costs $0.15 per million input tokens and Free per million output tokens. Prices last verified 4 Oct 2026.

### Is Gemini Embedding 001 open source?

No. Gemini Embedding 001 is a proprietary model; Google DeepMind has not released its weights. It is available through APIs including Gemini API, Google Vertex AI, Vercel AI Gateway.

### When was Gemini Embedding 001 released?

Google DeepMind released Gemini Embedding 001 on 20 May 2025, according to models.dev.

### Which providers offer Gemini Embedding 001?

We track Gemini Embedding 001 on 3 providers: Gemini API, Google Vertex AI, Vercel AI Gateway.

---
Data licensed CC BY 4.0. Cite as: AI Model Directory, "Gemini Embedding 001", https://www.aimodel.directory/models/gemini-embedding-001