Qwen: Qwen3 235B A22B vs Z.ai: GLM 4.6

Qwen: Qwen3 235B A22B is cheaper than Z.ai: GLM 4.6 by about 25% on a blended (75% input / 25% output) basis. Qwen: Qwen3 235B A22B charges $0.230000 / $2.300000 per 1M input/output tokens; Z.ai: GLM 4.6 charges $0.600000 / $2.200000.

FieldQwen: Qwen3 235B A22BZ.ai: GLM 4.6
Input $/Mtok$0.230000$0.600000
Output $/Mtok$2.300000$2.200000
Cache read $/Mtok–$0.110000
Cache write $/Mtok––
Input $/Mchar0.069354n/a
Output $/Mchar0.693544n/a
Context window128000202752
Tokenizerqwen2–

Price history

Same workload, monthly cost

Illustrative monthly cost at 10M tokens/month, split per workload preset. Not a substitute for the calculator, which uses your own volume, cache and batch assumptions.

WorkloadQwen: Qwen3 235B A22BZ.ai: GLM 4.6
Coding agent$3.33$6.80
RAG chat$5.40$8.40
Summarisation (batch)$4.37$7.60
Classification$2.71$6.32
Chat assistant$10.58$12.40
Content generation$18.86$18.80

Cite this page

Qwen: Qwen3 235B A22B vs Z.ai: GLM 4.6: $0.230000/$2.300000 vs $0.600000/$2.200000 per 1M tokens, as of 2026-10-04 (Token Barometer, https://tokenbarometer.com/compare/alibaba_qwen3-235b-a22b-vs-zhipu_glm-4.6).

Qwen: Qwen3 235B A22B page | Z.ai: GLM 4.6 page | Open in calculator