Together AI
Llama 3.3 70B (Together) token cost
Meta Llama 3.3 70B on Together — compare flat rates with Groq’s faster path.

Published list rates
Input
$1.04
per 1M tokens
Output
$1.04
per 1M tokens
Context window
131,072
tokens
Pricing source: Together AI pricing
Estimate cost
Estimate from text
Calculate from a token total
Project monthly volume
Daily requests × tokens per request × 30 days.
Estimated cost
- Estimated tokens
- —
- Characters
- —
- Words
- —
- Input tokens
- —
- Output tokens
- —
- Input cost
- —
- Output cost
- —
- Total cost
- —
- Combined price per 1M tokens
- —
- Estimated monthly cost
- —
- Estimated daily cost
- —
Cost as volume grows
Read down to see Llama 3.3 70B (Together) get more expensive as tokens grow. The last column is a typical mix: 70% input, 30% output.
| Tokens | Input only | Output only | Mix (70% in / 30% out) |
|---|---|---|---|
| 1 | $0.000001 | $0.000001 | $0.000001 |
| 10 | $0.00001 | $0.00001 | $0.00001 |
| 100 | $0.000104 | $0.000104 | $0.000104 |
| 200 | $0.000208 | $0.000208 | $0.000208 |
| 300 | $0.000312 | $0.000312 | $0.000312 |
| 400 | $0.000416 | $0.000416 | $0.000416 |
| 500 | $0.00052 | $0.00052 | $0.00052 |
| 1,000 | $0.00104 | $0.00104 | $0.00104 |
| 2,000 | $0.00208 | $0.00208 | $0.00208 |
| 5,000 | $0.0052 | $0.0052 | $0.0052 |
| 10,000 | $0.0104 | $0.0104 | $0.0104 |
| 25,000 | $0.026 | $0.026 | $0.026 |
| 50,000 | $0.052 | $0.052 | $0.052 |
| 100,000 | $0.104 | $0.104 | $0.104 |
| 250,000 | $0.26 | $0.26 | $0.26 |
| 500,000 | $0.52 | $0.52 | $0.52 |
| 1,000,000 | $1.04 | $1.04 | $1.04 |
| 5,000,000 | $5.20 | $5.20 | $5.20 |
| 10,000,000 | $10.40 | $10.40 | $10.40 |
| 50,000,000 | $52.00 | $52.00 | $52.00 |
| 100,000,000 | $104.00 | $104.00 | $104.00 |
| 500,000,000 | $520.00 | $520.00 | $520.00 |
| 1,000,000,000 | $1,040.00 | $1,040.00 | $1,040.00 |
| 10,000,000,000 | $10,400.00 | $10,400.00 | $10,400.00 |
| 100,000,000,000 | $104,000.00 | $104,000.00 | $104,000.00 |
Tokens for a fixed budget
Tokens Llama 3.3 70B (Together) can process for a set budget — send-only, reply-only, and a 70/30 mix.
| Budget | Input tokens | Output tokens | Mixed tokens (70/30) |
|---|---|---|---|
| $0.01 | 9,615 | 9,615 | 9,615 |
| $0.1 | 96,154 | 96,154 | 96,154 |
| $0.5 | 480,769 | 480,769 | 480,769 |
| $1.00 | 961,538 | 961,538 | 961,538 |
| $5.00 | 4.81M | 4.81M | 4.81M |
| $10.00 | 9.62M | 9.62M | 9.62M |
| $20.00 | 19.2M | 19.2M | 19.2M |
| $25.00 | 24M | 24M | 24M |
| $30.00 | 28.8M | 28.8M | 28.8M |
| $50.00 | 48.1M | 48.1M | 48.1M |
| $100.00 | 96.2M | 96.2M | 96.2M |
| $250.00 | 240.4M | 240.4M | 240.4M |
| $500.00 | 480.8M | 480.8M | 480.8M |
| $1,000.00 | 961.5M | 961.5M | 961.5M |
| $5,000.00 | 4.81B | 4.81B | 4.81B |
| $10,000.00 | 9.62B | 9.62B | 9.62B |
At these rates
- A workload with 70% input and 30% output costs $1.04 per 1 million total tokens and $104.00 per 100 million.
- At the same token count, output costs 1 times the input rate. Response length therefore deserves its own budget limit.
- Filling the 131,072-token context window with input alone would cost about $0.13631, before any output.
These figures use public list prices. Cache eligibility, batch or priority modes, long-context surcharges, discounts, and regional billing can change the invoice.