Groq
GPT OSS 20B (Groq) token cost
Smaller GPT OSS 20B on Groq for cheap, very fast completions and tool loops.

Published list rates
Input
$0.075
per 1M tokens
Output
$0.3
per 1M tokens
Context window
131,072
tokens
Pricing source: Groq pricing
Estimate cost
Estimate from text
Calculate from a token total
Project monthly volume
Daily requests × tokens per request × 30 days.
Estimated cost
- Estimated tokens
- —
- Characters
- —
- Words
- —
- Input tokens
- —
- Output tokens
- —
- Input cost
- —
- Output cost
- —
- Total cost
- —
- Combined price per 1M tokens
- —
- Estimated monthly cost
- —
- Estimated daily cost
- —
Cost as volume grows
GPT OSS 20B (Groq) at rising volume: prompt-only, reply-only, and a mixed chat (70% in, 30% out).
| Tokens | Input only | Output only | Mix (70% in / 30% out) |
|---|---|---|---|
| 1 | $0.000000075 | $0.0000003 | $0.0000001425 |
| 10 | $0.00000075 | $0.000003 | $0.000001 |
| 100 | $0.000008 | $0.00003 | $0.000014 |
| 200 | $0.000015 | $0.00006 | $0.000029 |
| 300 | $0.000022 | $0.00009 | $0.000043 |
| 400 | $0.00003 | $0.00012 | $0.000057 |
| 500 | $0.000038 | $0.00015 | $0.000071 |
| 1,000 | $0.000075 | $0.0003 | $0.000143 |
| 2,000 | $0.00015 | $0.0006 | $0.000285 |
| 5,000 | $0.000375 | $0.0015 | $0.000713 |
| 10,000 | $0.00075 | $0.003 | $0.001425 |
| 25,000 | $0.001875 | $0.0075 | $0.003562 |
| 50,000 | $0.00375 | $0.015 | $0.007125 |
| 100,000 | $0.0075 | $0.03 | $0.01425 |
| 250,000 | $0.01875 | $0.075 | $0.03563 |
| 500,000 | $0.0375 | $0.15 | $0.07125 |
| 1,000,000 | $0.075 | $0.3 | $0.1425 |
| 5,000,000 | $0.375 | $1.50 | $0.7125 |
| 10,000,000 | $0.75 | $3.00 | $1.425 |
| 50,000,000 | $3.75 | $15.00 | $7.125 |
| 100,000,000 | $7.50 | $30.00 | $14.25 |
| 500,000,000 | $37.50 | $150.00 | $71.25 |
| 1,000,000,000 | $75.00 | $300.00 | $142.50 |
| 10,000,000,000 | $750.00 | $3,000.00 | $1,425.00 |
| 100,000,000,000 | $7,500.00 | $30,000.00 | $14,250.00 |
Tokens for a fixed budget
How far a fixed spend goes on GPT OSS 20B (Groq) at these list rates.
| Budget | Input tokens | Output tokens | Mixed tokens (70/30) |
|---|---|---|---|
| $0.01 | 133,333 | 33,333 | 70,175 |
| $0.1 | 1.33M | 333,333 | 701,754 |
| $0.5 | 6.67M | 1.67M | 3.51M |
| $1.00 | 13.3M | 3.33M | 7.02M |
| $5.00 | 66.7M | 16.7M | 35.1M |
| $10.00 | 133.3M | 33.3M | 70.2M |
| $20.00 | 266.7M | 66.7M | 140.4M |
| $25.00 | 333.3M | 83.3M | 175.4M |
| $30.00 | 400M | 100M | 210.5M |
| $50.00 | 666.7M | 166.7M | 350.9M |
| $100.00 | 1.33B | 333.3M | 701.8M |
| $250.00 | 3.33B | 833.3M | 1.75B |
| $500.00 | 6.67B | 1.67B | 3.51B |
| $1,000.00 | 13.3B | 3.33B | 7.02B |
| $5,000.00 | 66.7B | 16.7B | 35.1B |
| $10,000.00 | 133.3B | 33.3B | 70.2B |
At these rates
- A workload with 70% input and 30% output costs $0.1425 per 1 million total tokens and $14.25 per 100 million.
- At the same token count, output costs 4 times the input rate. Response length therefore deserves its own budget limit.
- Filling the 131,072-token context window with input alone would cost about $0.00983, before any output.
These figures use public list prices. Cache eligibility, batch or priority modes, long-context surcharges, discounts, and regional billing can change the invoice.