Google: Gemini 2.5 Flash Lite vs Meta: Llama 3.3 70B Instruct

Meta: Llama 3.3 70B Instruct is cheaper than Google: Gemini 2.5 Flash Lite by about 100% on a blended (75% input / 25% output) basis. Google: Gemini 2.5 Flash Lite charges $0.100000 / $0.400000 per 1M input/output tokens; Meta: Llama 3.3 70B Instruct charges $0.000000 / $0.000000.

FieldGoogle: Gemini 2.5 Flash LiteMeta: Llama 3.3 70B Instruct
Input $/Mtok$0.100000$0.000000
Output $/Mtok$0.400000$0.000000
Cache read $/Mtok$0.010000–
Cache write $/Mtok––
Input $/Mcharn/a0.000000
Output $/Mcharn/a0.000000
Context window1048576128000
Tokenizergeminillama3

Price history

Same workload, monthly cost

Illustrative monthly cost at 10M tokens/month, split per workload preset. Not a substitute for the calculator, which uses your own volume, cache and batch assumptions.

WorkloadGoogle: Gemini 2.5 Flash LiteMeta: Llama 3.3 70B Instruct
Coding agent$1.15$0.00
RAG chat$1.45$0.00
Summarisation (batch)$1.30$0.00
Classification$1.06$0.00
Chat assistant$2.20$0.00
Content generation$3.40$0.00

Cite this page

Google: Gemini 2.5 Flash Lite vs Meta: Llama 3.3 70B Instruct: $0.100000/$0.400000 vs $0.000000/$0.000000 per 1M tokens, as of 2026-10-04 (Token Barometer, https://tokenbarometer.com/compare/google_gemini-2.5-flash-lite-vs-meta_llama-3.3-70b-instruct).

Google: Gemini 2.5 Flash Lite page | Meta: Llama 3.3 70B Instruct page | Open in calculator