llama-3.1-70b-instruct API pricing
Blended at 3:1 input to output: $0.400 per million tokens.
OpenRouter charges $0.400 per million input tokens and $0.400 per million output tokens. 2 of the models nearest it in price are cheaper on a blended basis.
Against the models nearest in price
Blended at a 3:1 input to output ratio.
| Source | ||||||
|---|---|---|---|---|---|---|
| llama-3.1-70b-instruct This model | OpenRouter | $0.400 | $0.400 | 131,072 | $0.400 | ↗ |
| Llama-4-Maverick-17B-128E-Instruct-FP8 | Together AI | $0.270 | $0.850 | — | $0.415 | ↗ |
| deepseek-chat-v3-0324 | OpenRouter | $0.250 | $1.00 | 65,536 | $0.438 | ↗ |
| gpt-5.6-luna | OpenAI | $0.200 | $1.20 | 922,000 | $0.450 | ↗ |
| deepseek-chat-v3.1 | OpenRouter | $0.200 | $0.800 | 163,840 | $0.350 | ↗ |
| deepseek-v3.1-terminus | OpenRouter | $0.270 | $1.00 | 163,840 | $0.453 | ↗ |
| gpt-5.4-nano | OpenAI | $0.200 | $1.25 | 272,000 | $0.463 | ↗ |
| llama-4-maverick | OpenRouter | $0.200 | $0.696 | 1,048,576 | $0.324 | ↗ |
| deepseek-v3 | DeepSeek | $0.270 | $1.10 | 65,536 | $0.477 | ↗ |
Last checked . Methodology
Questions
How much does llama-3.1-70b-instruct cost per million tokens?
$0.400 per million input tokens and $0.400 per million output tokens. At a 3:1 input to output mix that blends to $0.400. Checked 12 Sept 2026.
What would llama-3.1-70b-instruct cost for a real workload?
A thousand requests a day at 2,000 input and 600 output tokens each comes to $31.20 a month.
What is the context window for llama-3.1-70b-instruct?
131,072 tokens of input, with up to 16,384 tokens of output.
Llama-4-Maverick-17B-128E-Instruct-FP8 · deepseek-chat-v3-0324 · gpt-5.6-luna · deepseek-chat-v3.1 · deepseek-v3.1-terminus · All OpenRouter models · Cost calculator