←Model pricing

Kimi vs Zhipu API pricing

Checked against the official price pages on 2026-10-08


In short

At 100M input and 20M output tokens a month, the cheapest model from each is kimi-k2.7-code (about ¥1,190 a month) and GLM-5.3-Flash (about ¥136 a month). Zhipu's entry model is cheaper; Kimi's costs about 8.8 times as much.

The most expensive from each is kimi-k3 (about ¥4,000 a month) and GLM-5.3 (about ¥1,360 a month). Price says nothing about whether the models are equally capable; weigh it with benchmarks and context length.

For your own usage: API cost calculator

All models

CNY per million tokens; time-of-day pricing is written peak / off-peak. Monthly cost assumes 100M input and 20M output tokens a month, no cache hits and half of calls at peak hours where that applies, cheapest first.

ModelInputCache hitOutputMonthly
ZhipuGLM-5.3-Flash 0.80.232.8136
Kimikimi-k2.7-code 6.51.3271,190
Kimikimi-k2.6 6.51.1271,190
ZhipuGLM-5.3 82281,360
Kimikimi-k3 2021004,000

Price changes at both vendors

  1. 2026-07-16

    Kimi K3 launched (≈2.8T parameters, 1M-token context). International list price: $3 input / $15 output per million tokens (our Aug 27 issue). The China site now lists ¥20 input / ¥100 output. source · our coverage

  2. 2026-09-30

    Kimi K3.1 appeared on the API platform; the official price page does not list it yet. source · our coverage

Other comparisons