Qwen vs Zhipu API pricing
Checked against the official price pages on 2026-10-08
In short
At 100M input and 20M output tokens a month, the cheapest model from each is qwen3.7-plus (about ¥360 a month) and GLM-5.3-Flash (about ¥136 a month). Zhipu's entry model is cheaper; Qwen's costs about 2.6 times as much.
The most expensive from each is qwen3.8-max (about ¥1,920 a month) and GLM-5.3 (about ¥1,360 a month). Price says nothing about whether the models are equally capable; weigh it with benchmarks and context length.
For your own usage: API cost calculator
All models
CNY per million tokens; time-of-day pricing is written peak / off-peak. Monthly cost assumes 100M input and 20M output tokens a month, no cache hits and half of calls at peak hours where that applies, cheapest first.
| Model | Input | Cache hit | Output | Monthly |
|---|---|---|---|---|
| ZhipuGLM-5.3-Flash | 0.8 | 0.23 | 2.8 | 136 |
| Qwenqwen3.7-plus | 2 | — | 8 | 360 |
| ZhipuGLM-5.3 | 8 | 2 | 28 | 1,360 |
| Qwenqwen3.8-max | 12 | — | 36 | 1,920 |