Google: Gemini 2.5 Flash vs Meta: Llama 4 Scout

Meta: Llama 4 Scout is cheaper than Google: Gemini 2.5 Flash by about 93% on a blended (75% input / 25% output) basis. Google: Gemini 2.5 Flash charges $0.300000 / $2.500000 per 1M input/output tokens; Meta: Llama 4 Scout charges $0.050000 / $0.100000.

FieldGoogle: Gemini 2.5 FlashMeta: Llama 4 Scout
Input $/Mtok$0.300000$0.050000
Output $/Mtok$2.500000$0.100000
Cache read $/Mtok$0.030000–
Cache write $/Mtok––
Input $/Mcharn/a0.015047
Output $/Mcharn/a0.030093
Context window1048576131072
Tokenizergeminillama3

Price history

Same workload, monthly cost

Illustrative monthly cost at 10M tokens/month, split per workload preset. Not a substitute for the calculator, which uses your own volume, cache and batch assumptions.

WorkloadGoogle: Gemini 2.5 FlashMeta: Llama 4 Scout
Coding agent$4.10$0.53
RAG chat$6.30$0.58
Summarisation (batch)$5.20$0.55
Classification$3.44$0.51
Chat assistant$11.80$0.70
Content generation$20.60$0.90

Cite this page

Google: Gemini 2.5 Flash vs Meta: Llama 4 Scout: $0.300000/$2.500000 vs $0.050000/$0.100000 per 1M tokens, as of 2026-10-04 (Token Barometer, https://tokenbarometer.com/compare/google_gemini-2.5-flash-vs-meta_llama-4-scout).

Google: Gemini 2.5 Flash page | Meta: Llama 4 Scout page | Open in calculator