Llama 4 Maverick 17B 128E Instruct FP8 vs GPT-6 Astra (US)

Llama 4 Maverick 17B 128E Instruct FP8 is cheaper than GPT-6 Astra (US) by about 100% on a blended (75% input / 25% output) basis. Llama 4 Maverick 17B 128E Instruct FP8 charges $0.0500 / $0.10 per 1M input/output tokens; GPT-6 Astra (US) charges $10.00 / $50.00.

FieldLlama 4 Maverick 17B 128E Instruct FP8GPT-6 Astra (US)
Input $/Mtok$0.0500$10.00
Output $/Mtok$0.10$50.00
Cache read $/Mtok–$1.00
Cache write $/Mtok–$12.50
Input $/Mchar0.01502.82
Output $/Mchar0.030114.09
Context window5242881050000
Tokenizerllama3o200k
Confidence?aggregatorofficial
Sourcelitellmofficial-openai

Confidence: official = vendor pricing page · api = OpenRouter per-host · aggregator = community list, may lag · unreviewed = automatic extraction awaiting review. When sources disagree the most authoritative wins. Details: /methodology#confidence

Price history

Same workload, monthly cost

Illustrative monthly cost at 10M tokens/month, split per workload preset. Not a substitute for the calculator, which uses your own volume, cache and batch assumptions.

WorkloadLlama 4 Maverick 17B 128E Instruct FP8GPT-6 Astra (US)
Coding agent$0.53$120.00
RAG chat$0.58$160.00
Summarisation (batch)$0.55$140.00
Classification$0.51$108.00
Chat assistant$0.70$260.00
Content generation$0.90$420.00

Cite this page

Llama 4 Maverick 17B 128E Instruct FP8 vs GPT-6 Astra (US): $0.0500/$0.10 vs $10.00/$50.00 per 1M tokens as of 2026-10-04. Token Barometer by Dejan Strbad, https://tokenbarometer.com, data CC BY 4.0. https://tokenbarometer.com/compare/meta_llama-4-maverick-vs-openai_gpt-6-astra

Llama 4 Maverick 17B 128E Instruct FP8 page | GPT-6 Astra (US) page | Open in calculator