llama-4-maverick-17b-128e-instruct API pricing
Blended at 3:1 input to output: $0.300 per million tokens.
Groq charges $0.200 per million input tokens and $0.600 per million output tokens. 3 of the models nearest it in price are cheaper on a blended basis.
Against the models nearest in price
Blended at a 3:1 input to output ratio.
| Source | ||||||
|---|---|---|---|---|---|---|
| llama-4-maverick-17b-128e-instruct This model | Groq | $0.200 | $0.600 | 131,072 | $0.300 | ↗ |
| deepseek-v3.2 | OpenRouter | $0.269 | $0.400 | 163,840 | $0.302 | ↗ |
| deepseek-v3.2-exp | OpenRouter | $0.270 | $0.410 | 163,840 | $0.305 | ↗ |
| deepseek-chat | DeepSeek | $0.280 | $0.420 | 131,072 | $0.315 | ↗ |
| deepseek-reasoner | DeepSeek | $0.280 | $0.420 | 131,072 | $0.315 | ↗ |
| Llama-4-Scout-17B-16E-Instruct | Together AI | $0.180 | $0.590 | — | $0.282 | ↗ |
| llama-4-maverick | OpenRouter | $0.200 | $0.696 | 1,048,576 | $0.324 | ↗ |
| gpt-4o-mini | OpenAI | $0.150 | $0.600 | 128,000 | $0.263 | ↗ |
| gpt-4o-mini-audio-preview | OpenAI | $0.150 | $0.600 | 128,000 | $0.263 | ↗ |
Last checked . Methodology
Questions
How much does llama-4-maverick-17b-128e-instruct cost per million tokens?
$0.200 per million input tokens and $0.600 per million output tokens. At a 3:1 input to output mix that blends to $0.300. Checked 12 Sept 2026.
What would llama-4-maverick-17b-128e-instruct cost for a real workload?
A thousand requests a day at 2,000 input and 600 output tokens each comes to $22.80 a month.
What is the context window for llama-4-maverick-17b-128e-instruct?
131,072 tokens of input, with up to 8,192 tokens of output.
deepseek-v3.2 · deepseek-v3.2-exp · deepseek-chat · deepseek-reasoner · Llama-4-Scout-17B-16E-Instruct · All Groq models · Cost calculator