Meta: Llama 3.3 70B Instruct vs Mistral Small (latest)

Meta: Llama 3.3 70B Instruct is cheaper than Mistral Small (latest) by about 100% on a blended (75% input / 25% output) basis. Meta: Llama 3.3 70B Instruct charges $0.000000 / $0.000000 per 1M input/output tokens; Mistral Small (latest) charges $0.150000 / $0.600000.

FieldMeta: Llama 3.3 70B InstructMistral Small (latest)
Input $/Mtok$0.000000$0.150000
Output $/Mtok$0.000000$0.600000
Cache read $/Mtok–$0.015000
Cache write $/Mtok––
Input $/Mchar0.0000000.053343
Output $/Mchar0.0000000.213371
Context window128000256000
Tokenizerllama3mistral

Price history

Same workload, monthly cost

Illustrative monthly cost at 10M tokens/month, split per workload preset. Not a substitute for the calculator, which uses your own volume, cache and batch assumptions.

WorkloadMeta: Llama 3.3 70B InstructMistral Small (latest)
Coding agent$0.00$1.73
RAG chat$0.00$2.17
Summarisation (batch)$0.00$1.95
Classification$0.00$1.59
Chat assistant$0.00$3.30
Content generation$0.00$5.10

Cite this page

Meta: Llama 3.3 70B Instruct vs Mistral Small (latest): $0.000000/$0.000000 vs $0.150000/$0.600000 per 1M tokens, as of 2026-10-04 (Token Barometer, https://tokenbarometer.com/compare/meta_llama-3.3-70b-instruct-vs-mistral_mistral-small-2503).

Meta: Llama 3.3 70B Instruct page | Mistral Small (latest) page | Open in calculator