Llama-3.3-70B-Instruct-Turbo API pricing
Blended at 3:1 input to output: $1.04 per million tokens.
Together AI charges $1.04 per million input tokens and $1.04 per million output tokens. 7 of the models nearest it in price are cheaper on a blended basis.
Against the models nearest in price
Blended at a 3:1 input to output ratio.
| Source | ||||||
|---|---|---|---|---|---|---|
| Llama-3.3-70B-Instruct-Turbo This model | Together AI | $1.04 | $1.04 | 131,072 | $1.04 | ↗ |
| deepseek-v3-0324 | Fireworks AI | $0.900 | $0.900 | 163,840 | $0.900 | ↗ |
| mistral-large-3-fp8 | Fireworks AI | $1.20 | $1.20 | 256,000 | $1.20 | ↗ |
| DeepSeek-V3.1 | Together AI | $0.600 | $1.70 | 128,000 | $0.875 | ↗ |
| gemini-2.5-flash | Google AI | $0.300 | $2.50 | 1,048,576 | $0.850 | ↗ |
| gemini-2.5-flash-preview-09-2025 | Google AI | $0.300 | $2.50 | 1,048,576 | $0.850 | ↗ |
| deepseek-v3p1 | Fireworks AI | $0.560 | $1.68 | 128,000 | $0.840 | ↗ |
| deepseek-v3p1-terminus | Fireworks AI | $0.560 | $1.68 | 128,000 | $0.840 | ↗ |
| deepseek-v3p2 | Fireworks AI | $0.560 | $1.68 | 163,840 | $0.840 | ↗ |
Last checked . Methodology
Questions
How much does Llama-3.3-70B-Instruct-Turbo cost per million tokens?
$1.04 per million input tokens and $1.04 per million output tokens. At a 3:1 input to output mix that blends to $1.04. Checked 12 Sept 2026.
What would Llama-3.3-70B-Instruct-Turbo cost for a real workload?
A thousand requests a day at 2,000 input and 600 output tokens each comes to $81.12 a month.
What is the context window for Llama-3.3-70B-Instruct-Turbo?
131,072 tokens of input.
deepseek-v3-0324 · mistral-large-3-fp8 · DeepSeek-V3.1 · gemini-2.5-flash · gemini-2.5-flash-preview-09-2025 · All Together AI models · Cost calculator