GLM-5.3-Prime

Zhipu AI🇨🇳 China
active
Context window1000K tokens
Input / 1M tokens$2.8
Output / 1M tokens$8.8

Version History

5.3-Primeminor

GLM-5.3-Prime is a high-speed inference variant of GLM-5.3, delivering claimed 1.5-2x output throughput while retaining the same 1M-token context window and full capabilities. It targets coding and agentic workloads with always-on reasoning at low, high, or max effort.

Coverage