Inclusionai

Ling 3.1 Flash

Generally availableProprietary262K contextReasoning

Summary

Ling 3.1 Flash is a proprietary language model from Inclusionai, released on 2 Oct 2026. It accepts text and generates text, with a 262K-token context window and up to 33K output tokens per response.

Efficient model for low-latency assistance, extraction, and routine automation

Last verified 4 Oct 2026. Every figure on this page links to its source.

Specifications

DeveloperInclusionai
Release date2 Oct 2026 [models.dev]
StatusGenerally available [models.dev]
WeightsProprietary (API only) [models.dev]
Context window262K (262,144 tokens) [models.dev]
Max output33K tokens [models.dev]
InputText [models.dev]
OutputText [models.dev]
Reasoning modeYes [models.dev]
Tool callingYes [models.dev]
Structured outputNo [models.dev]

Ling 3.1 Flash API pricing by provider

USD per million tokens. Sorted by input price.

ProviderInputOutputCached inputContextSource
OpenRouter
inclusionai/ling-3.1-flash
FreeFree–262KOpenRouter, 4 Oct 2026
Vercel AI Gateway
inclusionai/ling-3.1-flash
FreeFree–262Kmodels.dev, 4 Oct 2026
Vercel AI Gateway
inclusionai/ling-3.1-flash-free
FreeFree–262Kmodels.dev, 4 Oct 2026

Frequently asked questions

What is the context window of Ling 3.1 Flash?

Ling 3.1 Flash has a context window of 262,144 tokens (262K) and can generate up to 32,768 tokens in a single response, according to models.dev.

Is Ling 3.1 Flash open source?

No. Ling 3.1 Flash is a proprietary model; Inclusionai has not released its weights. It is available through APIs including OpenRouter, Vercel AI Gateway.

When was Ling 3.1 Flash released?

Inclusionai released Ling 3.1 Flash on 2 Oct 2026, according to models.dev.

Which providers offer Ling 3.1 Flash?

We track Ling 3.1 Flash on 2 providers: OpenRouter, Vercel AI Gateway.

Change history

No changes detected since we started tracking this model. We re-check every source daily.

Sources