Google: Gemini 2.5 Flash vs Meta: Llama 4 Scout
Meta: Llama 4 Scout is cheaper than Google: Gemini 2.5 Flash by about 93% on a blended (75% input / 25% output) basis. Google: Gemini 2.5 Flash charges $0.300000 / $2.500000 per 1M input/output tokens; Meta: Llama 4 Scout charges $0.050000 / $0.100000.
| Field | Google: Gemini 2.5 Flash | Meta: Llama 4 Scout |
|---|---|---|
| Input $/Mtok | $0.300000 | $0.050000 |
| Output $/Mtok | $2.500000 | $0.100000 |
| Cache read $/Mtok | $0.030000 | – |
| Cache write $/Mtok | – | – |
| Input $/Mchar | n/a | 0.015047 |
| Output $/Mchar | n/a | 0.030093 |
| Context window | 1048576 | 131072 |
| Tokenizer | gemini | llama3 |
Price history
Same workload, monthly cost
Illustrative monthly cost at 10M tokens/month, split per workload preset. Not a substitute for the calculator, which uses your own volume, cache and batch assumptions.
| Workload | Google: Gemini 2.5 Flash | Meta: Llama 4 Scout |
|---|---|---|
| Coding agent | $4.10 | $0.53 |
| RAG chat | $6.30 | $0.58 |
| Summarisation (batch) | $5.20 | $0.55 |
| Classification | $3.44 | $0.51 |
| Chat assistant | $11.80 | $0.70 |
| Content generation | $20.60 | $0.90 |
Cite this page
Google: Gemini 2.5 Flash vs Meta: Llama 4 Scout: $0.300000/$2.500000 vs $0.050000/$0.100000 per 1M tokens, as of 2026-10-04 (Token Barometer, https://tokenbarometer.com/compare/google_gemini-2.5-flash-vs-meta_llama-4-scout).
Google: Gemini 2.5 Flash page | Meta: Llama 4 Scout page | Open in calculator