llama-3.3-70b-versatile API pricing

Checked 12 Sept 2026 LLMs
$ per 1M input $0.590
$ per 1M output $0.790

Blended at 3:1 input to output: $0.640 per million tokens.

Groq charges $0.590 per million input tokens and $0.790 per million output tokens. 2 of the models nearest it in price are cheaper on a blended basis.

$0.590
$/1M input
$0.790
$/1M output
131,072
Context window
$49.62
1k req/day, monthly

Against the models nearest in price

Blended at a 3:1 input to output ratio.

Calculator →
llama-3.3-70b-versatile against comparable models — checked 12 Sept 2026
Source
llama-3.3-70b-versatile This model Groq $0.590 $0.790 131,072 $0.640

Last checked . Methodology

Questions

How much does llama-3.3-70b-versatile cost per million tokens?

$0.590 per million input tokens and $0.790 per million output tokens. At a 3:1 input to output mix that blends to $0.640. Checked 12 Sept 2026.

What would llama-3.3-70b-versatile cost for a real workload?

A thousand requests a day at 2,000 input and 600 output tokens each comes to $49.62 a month.

What is the context window for llama-3.3-70b-versatile?

131,072 tokens of input, with up to 32,768 tokens of output.

llama-3-70b-instruct · gpt-5-mini · gpt-5.1-codex-mini · gpt-4.1-mini · mistral-large-2512 · All Groq models · Cost calculator