Google: Gemini 2.5 Flash Lite vs Microsoft: Phi 4

Microsoft: Phi 4 is cheaper than Google: Gemini 2.5 Flash Lite by about 50% on a blended (75% input / 25% output) basis. Google: Gemini 2.5 Flash Lite charges $0.100000 / $0.400000 per 1M input/output tokens; Microsoft: Phi 4 charges $0.070000 / $0.140000.

FieldGoogle: Gemini 2.5 Flash LiteMicrosoft: Phi 4
Input $/Mtok$0.100000$0.070000
Output $/Mtok$0.400000$0.140000
Cache read $/Mtok$0.010000–
Cache write $/Mtok––
Input $/Mcharn/an/a
Output $/Mcharn/an/a
Context window104857616384
Tokenizergemini–

Price history

Same workload, monthly cost

Illustrative monthly cost at 10M tokens/month, split per workload preset. Not a substitute for the calculator, which uses your own volume, cache and batch assumptions.

WorkloadGoogle: Gemini 2.5 Flash LiteMicrosoft: Phi 4
Coding agent$1.15$0.74
RAG chat$1.45$0.81
Summarisation (batch)$1.30$0.77
Classification$1.06$0.71
Chat assistant$2.20$0.98
Content generation$3.40$1.26

Cite this page

Google: Gemini 2.5 Flash Lite vs Microsoft: Phi 4: $0.100000/$0.400000 vs $0.070000/$0.140000 per 1M tokens, as of 2026-10-04 (Token Barometer, https://tokenbarometer.com/compare/google_gemini-2.5-flash-lite-vs-microsoft_phi-4).

Google: Gemini 2.5 Flash Lite page | Microsoft: Phi 4 page | Open in calculator