Llama 3.3 Nemotron Super 49B v1.5 vs Grok 3 mini
Llama 3.3 Nemotron Super 49B v1.5 is cheaper than Grok 3 mini by about 65% on a blended (75% input / 25% output) basis. Llama 3.3 Nemotron Super 49B v1.5 charges $0.10 / $0.40 per 1M input/output tokens; Grok 3 mini charges $0.25 / $1.27.
| Field | Llama 3.3 Nemotron Super 49B v1.5 | Grok 3 mini |
|---|---|---|
| Input $/Mtok | $0.10 | $0.25 |
| Output $/Mtok | $0.40 | $1.27 |
| Cache read $/Mtok | – | – |
| Cache write $/Mtok | – | – |
| Input $/Mchar | 0.0301 | n/a |
| Output $/Mchar | 0.12 | n/a |
| Context window | 131072 | 131072 |
| Tokenizer | llama3 | – |
| Confidence? | aggregator | aggregator |
| Source | litellm | litellm |
Confidence: official = vendor pricing page · api = OpenRouter per-host · aggregator = community list, may lag · unreviewed = automatic extraction awaiting review. When sources disagree the most authoritative wins. Details: /methodology#confidence
Price history
Same workload, monthly cost
Illustrative monthly cost at 10M tokens/month, split per workload preset. Not a substitute for the calculator, which uses your own volume, cache and batch assumptions.
| Workload | Llama 3.3 Nemotron Super 49B v1.5 | Grok 3 mini |
|---|---|---|
| Coding agent | $1.15 | $3.01 |
| RAG chat | $1.45 | $4.03 |
| Summarisation (batch) | $1.30 | $3.52 |
| Classification | $1.06 | $2.70 |
| Chat assistant | $2.20 | $6.58 |
| Content generation | $3.40 | $10.66 |
Cite this page
Llama 3.3 Nemotron Super 49B v1.5 vs Grok 3 mini: $0.10/$0.40 vs $0.25/$1.27 per 1M tokens as of 2026-10-04. Token Barometer by Dejan Strbad, https://tokenbarometer.com, data CC BY 4.0. https://tokenbarometer.com/compare/nvidia_nemotron-super-49b-v1-vs-xai_grok-3-mini
Llama 3.3 Nemotron Super 49B v1.5 page | Grok 3 mini page | Open in calculator