llama-3.1-8b-instruct API pricing

Checked 12 Sept 2026 LLMs
$ per 1M input $0.050
$ per 1M output $0.080

Blended at 3:1 input to output: $0.058 per million tokens.

OpenRouter charges $0.050 per million input tokens and $0.080 per million output tokens.

$0.050
$/1M input
$0.080
$/1M output
$0.025
$/1M cached in
131,072
Context window
$4.44
1k req/day, monthly

Against the models nearest in price

Blended at a 3:1 input to output ratio.

Calculator →
llama-3.1-8b-instruct against comparable models — checked 12 Sept 2026
Source
llama-3.1-8b-instruct This model OpenRouter $0.050 $0.080 131,072 $0.058

Last checked . Methodology

Questions

How much does llama-3.1-8b-instruct cost per million tokens?

$0.050 per million input tokens and $0.080 per million output tokens, with cached input at $0.025. At a 3:1 input to output mix that blends to $0.058. Checked 12 Sept 2026.

What would llama-3.1-8b-instruct cost for a real workload?

A thousand requests a day at 2,000 input and 600 output tokens each comes to $4.44 a month.

What is the context window for llama-3.1-8b-instruct?

131,072 tokens of input, with up to 117,964 tokens of output.

llama-3.1-8b-instant · llama-3.2-1b-instruct · llama-3.2-3b-instruct · gemini-2.0-flash-lite · gemini-2.0-flash-lite-001 · All OpenRouter models · Cost calculator