Sign InOpen Brain
VercelEngineering PostOfficial Source

GLM 5.3 now available on AI Gateway

GLM 5.3 is available through Vercel AI Gateway for coding agents, retaining a 1M-token context window while claiming better long-horizon engineering with fewer output tokens.

Vercel · Aug 18, 2026
Open Source Open MarkdownOpen JSON
Source Summary

Z.ai’s **GLM 5.3** is now on AI Gateway under zai/glm-5.3. It retains a **1M-token context window** and **128K-token maximum output**, with function calling, structured output, streaming, and context caching.

Practical Implication

Builders can route Claude Code, Codex, Cursor, OpenCode, Pi, and other agents to the model through Gateway. Evaluate it on long-running repository work and security tasks, tracking output-token use against GLM 5.2 at matched effort.

Agent-Ready Context
Z.ai’s **GLM 5.3** is now on AI Gateway under zai/glm-5.3. It retains a **1M-token context window** and **128K-token maximum output**, with function calling, structured output, streaming, and context caching.

Builders can route Claude Code, Codex, Cursor, OpenCode, Pi, and other agents to the model through Gateway. Evaluate it on long-running repository work and security tasks, tracking output-token use against GLM 5.2 at matched effort.

The engineering and vulnerability improvements are provider-reported, and the material gives no scores. The context and output limits are unchanged from GLM 5.2, so the release is about behavior and efficiency rather than larger capacity.
Connected Context · Feed7 Judgment

GLM 5.3 adds another repository-scale and security-oriented route without expanding GLM 5.2’s capacity envelope or supplying comparative scores. Its practical value is therefore a controlled behavioral and token-efficiency trial through existing agent clients, reinforcing that context size and provider claims cannot substitute for matched workload evaluation.

Context Map
modelcodingsecurity#model-selection#reasoning#coding-agents
Uncertainty
The engineering and vulnerability improvements are provider-reported, and the material gives no scores. The context and output limits are unchanged from GLM 5.2, so the release is about behavior and efficiency rather than larger capacity.