SIDE BY SIDE
Compare selected AI models side by side.
Compare uses, prices, speed evidence, and version details for up to three models. Language, video, decision models and historical evaluations are shown separately.
Language & coding
| What matters | Qwen3.8-Max ↗Alibaba Cloud |
|---|---|
| Use cases | Coding · Agents & tools · Documents & writing · Visual understanding |
| Model score | Not reviewed |
| Output / 1M tokens · USD | $2 |
| Published rates · USD | $2 input / $6 outputPer 1M native tokens |
| Cached input / 1M | $0.25 |
| Cache write / 1M | $2.5 |
| Pricing conditions | Model Studio · Singapore / InternationalCached input shows implicit caching. Explicit cache read is $0.17/MTok. Regional and promotional rates differ.Qwen3.8-Max · Singapore pricing ↗ |
| Output speed · tokens/s | Not benchmarkedNo speed classification verified for this version and serving tier. This does not mean it is slow. |
| Context window | 1,000,000 |
| Model ID | qwen3.8-max |
| Version notes | The recorded rates are for the Singapore deployment, not a conversion of mainland China prices. |
| Official documentation | Qwen3.8-Max · Singapore pricing ↗Checked 2026-10-05 |
Language prices are USD per million tokens; input and output are billed separately. Video samples use 5s of 768P output where verified. Extra tools, reasoning, cache writes, input media, storage, and taxes can change the total. (fast) marks a published median output speed above 150 tokens/s in the linked test settings.