Mistral Small (latest) vs Llama 3.3 Nemotron Super 49B v1.5

Llama 3.3 Nemotron Super 49B v1.5 is cheaper than Mistral Small (latest) by about 100% on a blended (75% input / 25% output) basis. Mistral Small (latest) charges $0.150000 / $0.600000 per 1M input/output tokens; Llama 3.3 Nemotron Super 49B v1.5 charges $0.000000 / $0.000000.

FieldMistral Small (latest)Llama 3.3 Nemotron Super 49B v1.5
Input $/Mtok$0.150000$0.000000
Output $/Mtok$0.600000$0.000000
Cache read $/Mtok$0.015000–
Cache write $/Mtok––
Input $/Mchar0.0533430.000000
Output $/Mchar0.2133710.000000
Context window256000131072
Tokenizermistralllama3

Price history

Same workload, monthly cost

Illustrative monthly cost at 10M tokens/month, split per workload preset. Not a substitute for the calculator, which uses your own volume, cache and batch assumptions.

WorkloadMistral Small (latest)Llama 3.3 Nemotron Super 49B v1.5
Coding agent$1.73$0.00
RAG chat$2.17$0.00
Summarisation (batch)$1.95$0.00
Classification$1.59$0.00
Chat assistant$3.30$0.00
Content generation$5.10$0.00

Cite this page

Mistral Small (latest) vs Llama 3.3 Nemotron Super 49B v1.5: $0.150000/$0.600000 vs $0.000000/$0.000000 per 1M tokens, as of 2026-10-04 (Token Barometer, https://tokenbarometer.com/compare/mistral_mistral-small-2503-vs-nvidia_nemotron-super-49b-v1).

Mistral Small (latest) page | Llama 3.3 Nemotron Super 49B v1.5 page | Open in calculator