llama-4-maverick-17b-128e-instruct API pricing

Checked 12 Sept 2026 LLMs
$ per 1M input $0.200
$ per 1M output $0.600

Blended at 3:1 input to output: $0.300 per million tokens.

Groq charges $0.200 per million input tokens and $0.600 per million output tokens. 3 of the models nearest it in price are cheaper on a blended basis.

$0.200
$/1M input
$0.600
$/1M output
131,072
Context window
$22.80
1k req/day, monthly

Against the models nearest in price

Blended at a 3:1 input to output ratio.

Calculator →
llama-4-maverick-17b-128e-instruct against comparable models — checked 12 Sept 2026
Source
llama-4-maverick-17b-128e-instruct This model Groq $0.200 $0.600 131,072 $0.300

Last checked . Methodology

Questions

How much does llama-4-maverick-17b-128e-instruct cost per million tokens?

$0.200 per million input tokens and $0.600 per million output tokens. At a 3:1 input to output mix that blends to $0.300. Checked 12 Sept 2026.

What would llama-4-maverick-17b-128e-instruct cost for a real workload?

A thousand requests a day at 2,000 input and 600 output tokens each comes to $22.80 a month.

What is the context window for llama-4-maverick-17b-128e-instruct?

131,072 tokens of input, with up to 8,192 tokens of output.

deepseek-v3.2 · deepseek-v3.2-exp · deepseek-chat · deepseek-reasoner · Llama-4-Scout-17B-16E-Instruct · All Groq models · Cost calculator