Google: Gemini 2.5 Flash vs Z.ai: GLM 4.6

Google: Gemini 2.5 Flash is cheaper than Z.ai: GLM 4.6 by about 15% on a blended (75% input / 25% output) basis. Google: Gemini 2.5 Flash charges $0.300000 / $2.500000 per 1M input/output tokens; Z.ai: GLM 4.6 charges $0.600000 / $2.200000.

FieldGoogle: Gemini 2.5 FlashZ.ai: GLM 4.6
Input $/Mtok$0.300000$0.600000
Output $/Mtok$2.500000$2.200000
Cache read $/Mtok$0.030000$0.110000
Cache write $/Mtok––
Input $/Mcharn/an/a
Output $/Mcharn/an/a
Context window1048576202752
Tokenizergemini–

Price history

Same workload, monthly cost

Illustrative monthly cost at 10M tokens/month, split per workload preset. Not a substitute for the calculator, which uses your own volume, cache and batch assumptions.

WorkloadGoogle: Gemini 2.5 FlashZ.ai: GLM 4.6
Coding agent$4.10$6.80
RAG chat$6.30$8.40
Summarisation (batch)$5.20$7.60
Classification$3.44$6.32
Chat assistant$11.80$12.40
Content generation$20.60$18.80

Cite this page

Google: Gemini 2.5 Flash vs Z.ai: GLM 4.6: $0.300000/$2.500000 vs $0.600000/$2.200000 per 1M tokens, as of 2026-10-04 (Token Barometer, https://tokenbarometer.com/compare/google_gemini-2.5-flash-vs-zhipu_glm-4.6).

Google: Gemini 2.5 Flash page | Z.ai: GLM 4.6 page | Open in calculator