GLM-5.2 Turbo

Zhipu's faster GLM-5.2 variant tuned for long-running agent workflows

Paid Language
Visit Product Page →
202752 context tokens
Undisclosed parameters
Proprietary license
Aug 2026 released

Zhipu AI released GLM-5.2 Turbo on August 17, 2026 as a faster sibling to GLM-5.2, built specifically for agent-driven workloads rather than general chat or one-off queries. It carries a smaller context window than the base model, about 200,000 tokens versus GLM-5.2’s full 1-million-token capacity, in exchange for lower latency on the long execution chains typical of coding agents and automation tools. Zhipu says it improved complex instruction decomposition and tool use, along with stability across scheduled and persistent tasks that run for extended periods without human supervision.

Pricing sits above the base GLM-5.2 rate, at $1.20 per million input tokens and $4 per million output tokens, reflecting the tradeoff toward speed over raw context length. It’s aimed at the same agent-framework crowd already using GLM-5.2 and GLM-5.3, giving developers a cheaper-context, faster-response option when a task doesn’t need the full million-token window.