# MiMo-V2-Flash

> MiMo-V2-Flash is an open-weight language model from Xiaomi (MiMo), released on 16 Dec 2025. It accepts text and generates text, with a 262K-token context window and up to 66K output tokens per response. As of 4 Oct 2026, the lowest listed API price is $0.10 per million input tokens and $0.30 per million output tokens (Novita AI), across 2 providers we track. Xiaomi (MiMo) has deprecated this model.

Source page: https://www.aimodel.directory/models/mimo-v2-flash
Last verified: 4 Oct 2026

## Specifications

- **Developer:** Xiaomi (MiMo)
- **Release date:** 16 Dec 2025 (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/xiaomi/models/mimo-v2-flash.toml)
- **Status:** Deprecated (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/xiaomi/models/mimo-v2-flash.toml)
- **Weights:** Open weights (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/xiaomi/models/mimo-v2-flash.toml)
- **Context window:** 262K (262,144 tokens) (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/xiaomi/models/mimo-v2-flash.toml)
- **Max output:** 66K tokens (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/xiaomi/models/mimo-v2-flash.toml)
- **Knowledge cutoff:** 1 Dec 2024 (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/xiaomi/models/mimo-v2-flash.toml)
- **Input:** Text (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/xiaomi/models/mimo-v2-flash.toml)
- **Output:** Text (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/xiaomi/models/mimo-v2-flash.toml)
- **Reasoning mode:** Yes (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/xiaomi/models/mimo-v2-flash.toml)
- **Tool calling:** Yes (source: models.dev, https://github.com/sst/models.dev/blob/dev/providers/xiaomi/models/mimo-v2-flash.toml)

## API pricing (USD per 1M tokens)

| Provider | Input | Output | Cached input | Context |
|---|---|---|---|---|
| Novita AI | $0.10 | $0.30 | $0.30 | 262K |
| Hugging Face Inference Providers | $0.10 | $0.30 | – | 262K |

## FAQ

### What is the context window of MiMo-V2-Flash?

MiMo-V2-Flash has a context window of 262,144 tokens (262K) and can generate up to 65,536 tokens in a single response, according to models.dev.

### How much does the MiMo-V2-Flash API cost?

MiMo-V2-Flash is listed from $0.10 per million input tokens and $0.30 per million output tokens (Novita AI). Prices last verified 4 Oct 2026.

### Is MiMo-V2-Flash open source?

MiMo-V2-Flash is an open-weight model. Check the license terms before commercial use.

### When was MiMo-V2-Flash released?

Xiaomi (MiMo) released MiMo-V2-Flash on 16 Dec 2025, according to models.dev.

### Which providers offer MiMo-V2-Flash?

We track MiMo-V2-Flash on 2 providers: Novita AI, Hugging Face Inference Providers.

---
Data licensed CC BY 4.0. Cite as: AI Model Directory, "MiMo-V2-Flash", https://www.aimodel.directory/models/mimo-v2-flash