Llama 3.3 Nemotron Super 49B v1.5 vs Grok 3 mini

Llama 3.3 Nemotron Super 49B v1.5 is cheaper than Grok 3 mini by about 100% on a blended (75% input / 25% output) basis. Llama 3.3 Nemotron Super 49B v1.5 charges $0.000000 / $0.000000 per 1M input/output tokens; Grok 3 mini charges $0.250000 / $1.270000.

FieldLlama 3.3 Nemotron Super 49B v1.5Grok 3 mini
Input $/Mtok$0.000000$0.250000
Output $/Mtok$0.000000$1.270000
Cache read $/Mtok––
Cache write $/Mtok––
Input $/Mchar0.000000n/a
Output $/Mchar0.000000n/a
Context window131072131072
Tokenizerllama3–

Price history

Same workload, monthly cost

Illustrative monthly cost at 10M tokens/month, split per workload preset. Not a substitute for the calculator, which uses your own volume, cache and batch assumptions.

WorkloadLlama 3.3 Nemotron Super 49B v1.5Grok 3 mini
Coding agent$0.00$3.01
RAG chat$0.00$4.03
Summarisation (batch)$0.00$3.52
Classification$0.00$2.70
Chat assistant$0.00$6.58
Content generation$0.00$10.66

Cite this page

Llama 3.3 Nemotron Super 49B v1.5 vs Grok 3 mini: $0.000000/$0.000000 vs $0.250000/$1.270000 per 1M tokens, as of 2026-10-04 (Token Barometer, https://tokenbarometer.com/compare/nvidia_nemotron-super-49b-v1-vs-xai_grok-3-mini).

Llama 3.3 Nemotron Super 49B v1.5 page | Grok 3 mini page | Open in calculator