Qwen3 Coder 480B A35B Instruct vs Z.ai: GLM 4.6

Qwen3 Coder 480B A35B Instruct is cheaper than Z.ai: GLM 4.6 by about 100% on a blended (75% input / 25% output) basis. Qwen3 Coder 480B A35B Instruct charges $0.000000 / $0.000000 per 1M input/output tokens; Z.ai: GLM 4.6 charges $0.600000 / $2.200000.

FieldQwen3 Coder 480B A35B InstructZ.ai: GLM 4.6
Input $/Mtok$0.000000$0.600000
Output $/Mtok$0.000000$2.200000
Cache read $/Mtok–$0.110000
Cache write $/Mtok––
Input $/Mchar0.000000n/a
Output $/Mchar0.000000n/a
Context window262144202752
Tokenizerqwen2–

Price history

Same workload, monthly cost

Illustrative monthly cost at 10M tokens/month, split per workload preset. Not a substitute for the calculator, which uses your own volume, cache and batch assumptions.

WorkloadQwen3 Coder 480B A35B InstructZ.ai: GLM 4.6
Coding agent$0.00$6.80
RAG chat$0.00$8.40
Summarisation (batch)$0.00$7.60
Classification$0.00$6.32
Chat assistant$0.00$12.40
Content generation$0.00$18.80

Cite this page

Qwen3 Coder 480B A35B Instruct vs Z.ai: GLM 4.6: $0.000000/$0.000000 vs $0.600000/$2.200000 per 1M tokens, as of 2026-10-04 (Token Barometer, https://tokenbarometer.com/compare/alibaba_qwen3-coder-480b-a35b-vs-zhipu_glm-4.6).

Qwen3 Coder 480B A35B Instruct page | Z.ai: GLM 4.6 page | Open in calculator