Zhipu AI (Z.ai)

GLM 5.3 Prime

Generally available1M contextReasoning

Summary

GLM 5.3 Prime is a language model from Zhipu AI (Z.ai), released on 23 Sept 2026. It accepts text and generates text, with a 1M-token context window and up to 131K output tokens per response. As of 4 Oct 2026, the lowest listed API price is $2.80 per million input tokens and $8.80 per million output tokens (OpenRouter), across 1 provider we track.

GLM-5.3-Prime is the high-speed variant of Z.ai's GLM-5.3, inheriting its full capabilities while delivering 1.5–2× the output throughput through inference acceleration. It supports text input and output with a 1M-token...

Last verified 4 Oct 2026. Every figure on this page links to its source.

Specifications

DeveloperZhipu AI (Z.ai)
Release date23 Sept 2026 [OpenRouter]
StatusGenerally available
Context window1M (1,000,000 tokens) [OpenRouter]
Max output131K tokens [OpenRouter]
InputText [OpenRouter]
OutputText [OpenRouter]
Reasoning modeYes [OpenRouter]
Tool callingYes [OpenRouter]
Structured outputNo [OpenRouter]

GLM 5.3 Prime API pricing by provider

USD per million tokens. Sorted by input price.

ProviderInputOutputCached inputContextSource
OpenRouter
z-ai/glm-5.3-prime
$2.80$8.80$0.561MOpenRouter, 4 Oct 2026

Frequently asked questions

What is the context window of GLM 5.3 Prime?

GLM 5.3 Prime has a context window of 1,000,000 tokens (1M) and can generate up to 131,072 tokens in a single response, according to OpenRouter.

How much does the GLM 5.3 Prime API cost?

GLM 5.3 Prime is listed from $2.80 per million input tokens and $8.80 per million output tokens (OpenRouter). Prices last verified 4 Oct 2026.

When was GLM 5.3 Prime released?

Zhipu AI (Z.ai) released GLM 5.3 Prime on 23 Sept 2026, according to OpenRouter.

Change history

No changes detected since we started tracking this model. We re-check every source daily.

Sources