Google: Gemini 2.5 Flash Lite vs Llama 3.3 Nemotron Super 49B v1.5
Llama 3.3 Nemotron Super 49B v1.5 is cheaper than Google: Gemini 2.5 Flash Lite by about 100% on a blended (75% input / 25% output) basis. Google: Gemini 2.5 Flash Lite charges $0.100000 / $0.400000 per 1M input/output tokens; Llama 3.3 Nemotron Super 49B v1.5 charges $0.000000 / $0.000000.
| Field | Google: Gemini 2.5 Flash Lite | Llama 3.3 Nemotron Super 49B v1.5 |
|---|---|---|
| Input $/Mtok | $0.100000 | $0.000000 |
| Output $/Mtok | $0.400000 | $0.000000 |
| Cache read $/Mtok | $0.010000 | – |
| Cache write $/Mtok | – | – |
| Input $/Mchar | n/a | 0.000000 |
| Output $/Mchar | n/a | 0.000000 |
| Context window | 1048576 | 131072 |
| Tokenizer | gemini | llama3 |
Price history
Same workload, monthly cost
Illustrative monthly cost at 10M tokens/month, split per workload preset. Not a substitute for the calculator, which uses your own volume, cache and batch assumptions.
| Workload | Google: Gemini 2.5 Flash Lite | Llama 3.3 Nemotron Super 49B v1.5 |
|---|---|---|
| Coding agent | $1.15 | $0.00 |
| RAG chat | $1.45 | $0.00 |
| Summarisation (batch) | $1.30 | $0.00 |
| Classification | $1.06 | $0.00 |
| Chat assistant | $2.20 | $0.00 |
| Content generation | $3.40 | $0.00 |
Cite this page
Google: Gemini 2.5 Flash Lite vs Llama 3.3 Nemotron Super 49B v1.5: $0.100000/$0.400000 vs $0.000000/$0.000000 per 1M tokens, as of 2026-10-04 (Token Barometer, https://tokenbarometer.com/compare/google_gemini-2.5-flash-lite-vs-nvidia_nemotron-super-49b-v1).
Google: Gemini 2.5 Flash Lite page | Llama 3.3 Nemotron Super 49B v1.5 page | Open in calculator