Together AI

Llama 3.3 70B (Together) token cost

Meta Llama 3.3 70B on Together — compare flat rates with Groq’s faster path.

Published list rates

Input

$1.04

per 1M tokens

Output

$1.04

per 1M tokens

Context window

131,072

tokens

Pricing source: Together AI pricing

Estimate cost

Estimate from text

Calculate from a token total

Project monthly volume

Daily requests × tokens per request × 30 days.

Estimated cost

Estimated tokens
Characters
Words
Input tokens
Output tokens
Input cost
Output cost
Total cost
Combined price per 1M tokens
Estimated monthly cost
Estimated daily cost

Cost as volume grows

Read down to see Llama 3.3 70B (Together) get more expensive as tokens grow. The last column is a typical mix: 70% input, 30% output.

Tokens Input only Output only Mix (70% in / 30% out)
1 $0.000001 $0.000001 $0.000001
10 $0.00001 $0.00001 $0.00001
100 $0.000104 $0.000104 $0.000104
200 $0.000208 $0.000208 $0.000208
300 $0.000312 $0.000312 $0.000312
400 $0.000416 $0.000416 $0.000416
500 $0.00052 $0.00052 $0.00052
1,000 $0.00104 $0.00104 $0.00104
2,000 $0.00208 $0.00208 $0.00208
5,000 $0.0052 $0.0052 $0.0052
10,000 $0.0104 $0.0104 $0.0104
25,000 $0.026 $0.026 $0.026
50,000 $0.052 $0.052 $0.052
100,000 $0.104 $0.104 $0.104
250,000 $0.26 $0.26 $0.26
500,000 $0.52 $0.52 $0.52
1,000,000 $1.04 $1.04 $1.04
5,000,000 $5.20 $5.20 $5.20
10,000,000 $10.40 $10.40 $10.40
50,000,000 $52.00 $52.00 $52.00
100,000,000 $104.00 $104.00 $104.00
500,000,000 $520.00 $520.00 $520.00
1,000,000,000 $1,040.00 $1,040.00 $1,040.00
10,000,000,000 $10,400.00 $10,400.00 $10,400.00
100,000,000,000 $104,000.00 $104,000.00 $104,000.00

Tokens for a fixed budget

Tokens Llama 3.3 70B (Together) can process for a set budget — send-only, reply-only, and a 70/30 mix.

Budget Input tokens Output tokens Mixed tokens (70/30)
$0.01 9,615 9,615 9,615
$0.1 96,154 96,154 96,154
$0.5 480,769 480,769 480,769
$1.00 961,538 961,538 961,538
$5.00 4.81M 4.81M 4.81M
$10.00 9.62M 9.62M 9.62M
$20.00 19.2M 19.2M 19.2M
$25.00 24M 24M 24M
$30.00 28.8M 28.8M 28.8M
$50.00 48.1M 48.1M 48.1M
$100.00 96.2M 96.2M 96.2M
$250.00 240.4M 240.4M 240.4M
$500.00 480.8M 480.8M 480.8M
$1,000.00 961.5M 961.5M 961.5M
$5,000.00 4.81B 4.81B 4.81B
$10,000.00 9.62B 9.62B 9.62B

At these rates

  • A workload with 70% input and 30% output costs $1.04 per 1 million total tokens and $104.00 per 100 million.
  • At the same token count, output costs 1 times the input rate. Response length therefore deserves its own budget limit.
  • Filling the 131,072-token context window with input alone would cost about $0.13631, before any output.

Together AI

Open-model serverless inference, image, audio, and dedicated GPU endpoints.

All Together AI models

These figures use public list prices. Cache eligibility, batch or priority modes, long-context surcharges, discounts, and regional billing can change the invoice.

Related models