Llama 3.3 Nemotron Super 49B v1.5 vs Grok 3 mini

Llama 3.3 Nemotron Super 49B v1.5 is cheaper than Grok 3 mini by about 65% on a blended (75% input / 25% output) basis. Llama 3.3 Nemotron Super 49B v1.5 charges $0.10 / $0.40 per 1M input/output tokens; Grok 3 mini charges $0.25 / $1.27.

FieldLlama 3.3 Nemotron Super 49B v1.5Grok 3 mini
Input $/Mtok$0.10$0.25
Output $/Mtok$0.40$1.27
Cache read $/Mtok––
Cache write $/Mtok––
Input $/Mchar0.0301n/a
Output $/Mchar0.12n/a
Context window131072131072
Tokenizerllama3–
Confidence?aggregatoraggregator
Sourcelitellmlitellm

Confidence: official = vendor pricing page · api = OpenRouter per-host · aggregator = community list, may lag · unreviewed = automatic extraction awaiting review. When sources disagree the most authoritative wins. Details: /methodology#confidence

Price history

Same workload, monthly cost

Illustrative monthly cost at 10M tokens/month, split per workload preset. Not a substitute for the calculator, which uses your own volume, cache and batch assumptions.

WorkloadLlama 3.3 Nemotron Super 49B v1.5Grok 3 mini
Coding agent$1.15$3.01
RAG chat$1.45$4.03
Summarisation (batch)$1.30$3.52
Classification$1.06$2.70
Chat assistant$2.20$6.58
Content generation$3.40$10.66

Cite this page

Llama 3.3 Nemotron Super 49B v1.5 vs Grok 3 mini: $0.10/$0.40 vs $0.25/$1.27 per 1M tokens as of 2026-10-04. Token Barometer by Dejan Strbad, https://tokenbarometer.com, data CC BY 4.0. https://tokenbarometer.com/compare/nvidia_nemotron-super-49b-v1-vs-xai_grok-3-mini

Llama 3.3 Nemotron Super 49B v1.5 page | Grok 3 mini page | Open in calculator