SIDE BY SIDE
Compare selected AI models side by side.
Compare uses, prices, speed evidence, and version details for up to three models. Language, video, decision models and historical evaluations are shown separately.
Language & coding
| What matters | GLM-5.3 ↗Z.ai |
|---|---|
| Use cases | Coding · Agents & tools · Reasoning |
| Model score | 45GLM-5.3 · maxArtificial Analysis · GLM-5.3 · max ↗ |
| Output / 1M tokens · USD | $1.4 |
| Published rates · USD | $1.4 input / $4.4 outputPer 1M native tokens |
| Cached input / 1M | $0.26 |
| Cache write / 1M | Not verified |
| Pricing conditions | Z.ai API · listed USD ratesCached storage is listed as temporarily free. Future storage charges are not estimated here.Z.ai API pricing ↗ |
| Output speed · tokens/s | Not benchmarkedNo speed classification verified for this version and serving tier. This does not mean it is slow. |
| Context window | 1,000,000 |
| Model ID | glm-5.3 |
| Version notes | Reasoning is always enabled. Supports low, high, and max effort; image input is not supported. |
| Official documentation | GLM-5.3 model ↗Checked 2026-10-05 |
Language prices are USD per million tokens; input and output are billed separately. Video samples use 5s of 768P output where verified. Extra tools, reasoning, cache writes, input media, storage, and taxes can change the total. (fast) marks a published median output speed above 150 tokens/s in the linked test settings.