← Choose models

Language & coding

What mattersGLM-5.3 ↗Z.ai
Use casesCoding · Agents & tools · Reasoning
Model score45GLM-5.3 · maxArtificial Analysis · GLM-5.3 · max ↗
Output / 1M tokens · USD$1.4
Published rates · USD$1.4 input / $4.4 outputPer 1M native tokens
Cached input / 1M$0.26
Cache write / 1MNot verified
Pricing conditionsZ.ai API · listed USD ratesCached storage is listed as temporarily free. Future storage charges are not estimated here.Z.ai API pricing ↗
Output speed · tokens/sNot benchmarkedNo speed classification verified for this version and serving tier. This does not mean it is slow.
Context window1,000,000
Model IDglm-5.3
Version notesReasoning is always enabled. Supports low, high, and max effort; image input is not supported.
Official documentationGLM-5.3 model ↗Checked 2026-10-05
Language prices are USD per million tokens; input and output are billed separately. Video samples use 5s of 768P output where verified. Extra tools, reasoning, cache writes, input media, storage, and taxes can change the total. (fast) marks a published median output speed above 150 tokens/s in the linked test settings.