Llama 4 Maverick 17B 128E Instruct FP8 vs GPT-6 Astra (US)
Llama 4 Maverick 17B 128E Instruct FP8 is cheaper than GPT-6 Astra (US) by about 100% on a blended (75% input / 25% output) basis. Llama 4 Maverick 17B 128E Instruct FP8 charges $0.0500 / $0.10 per 1M input/output tokens; GPT-6 Astra (US) charges $10.00 / $50.00.
| Field | Llama 4 Maverick 17B 128E Instruct FP8 | GPT-6 Astra (US) |
|---|---|---|
| Input $/Mtok | $0.0500 | $10.00 |
| Output $/Mtok | $0.10 | $50.00 |
| Cache read $/Mtok | – | $1.00 |
| Cache write $/Mtok | – | $12.50 |
| Input $/Mchar | 0.0150 | 2.82 |
| Output $/Mchar | 0.0301 | 14.09 |
| Context window | 524288 | 1050000 |
| Tokenizer | llama3 | o200k |
| Confidence? | aggregator | official |
| Source | litellm | official-openai |
Confidence: official = vendor pricing page · api = OpenRouter per-host · aggregator = community list, may lag · unreviewed = automatic extraction awaiting review. When sources disagree the most authoritative wins. Details: /methodology#confidence
Price history
Same workload, monthly cost
Illustrative monthly cost at 10M tokens/month, split per workload preset. Not a substitute for the calculator, which uses your own volume, cache and batch assumptions.
| Workload | Llama 4 Maverick 17B 128E Instruct FP8 | GPT-6 Astra (US) |
|---|---|---|
| Coding agent | $0.53 | $120.00 |
| RAG chat | $0.58 | $160.00 |
| Summarisation (batch) | $0.55 | $140.00 |
| Classification | $0.51 | $108.00 |
| Chat assistant | $0.70 | $260.00 |
| Content generation | $0.90 | $420.00 |
Cite this page
Llama 4 Maverick 17B 128E Instruct FP8 vs GPT-6 Astra (US): $0.0500/$0.10 vs $10.00/$50.00 per 1M tokens as of 2026-10-04. Token Barometer by Dejan Strbad, https://tokenbarometer.com, data CC BY 4.0. https://tokenbarometer.com/compare/meta_llama-4-maverick-vs-openai_gpt-6-astra
Llama 4 Maverick 17B 128E Instruct FP8 page | GPT-6 Astra (US) page | Open in calculator