Groq

GPT OSS 20B (Groq) token cost

Smaller GPT OSS 20B on Groq for cheap, very fast completions and tool loops.

Published list rates

Input

$0.075

per 1M tokens

Output

$0.3

per 1M tokens

Context window

131,072

tokens

Pricing source: Groq pricing

Estimate cost

Estimate from text

Calculate from a token total

Project monthly volume

Daily requests × tokens per request × 30 days.

Estimated cost

Estimated tokens
Characters
Words
Input tokens
Output tokens
Input cost
Output cost
Total cost
Combined price per 1M tokens
Estimated monthly cost
Estimated daily cost

Cost as volume grows

GPT OSS 20B (Groq) at rising volume: prompt-only, reply-only, and a mixed chat (70% in, 30% out).

Tokens Input only Output only Mix (70% in / 30% out)
1 $0.000000075 $0.0000003 $0.0000001425
10 $0.00000075 $0.000003 $0.000001
100 $0.000008 $0.00003 $0.000014
200 $0.000015 $0.00006 $0.000029
300 $0.000022 $0.00009 $0.000043
400 $0.00003 $0.00012 $0.000057
500 $0.000038 $0.00015 $0.000071
1,000 $0.000075 $0.0003 $0.000143
2,000 $0.00015 $0.0006 $0.000285
5,000 $0.000375 $0.0015 $0.000713
10,000 $0.00075 $0.003 $0.001425
25,000 $0.001875 $0.0075 $0.003562
50,000 $0.00375 $0.015 $0.007125
100,000 $0.0075 $0.03 $0.01425
250,000 $0.01875 $0.075 $0.03563
500,000 $0.0375 $0.15 $0.07125
1,000,000 $0.075 $0.3 $0.1425
5,000,000 $0.375 $1.50 $0.7125
10,000,000 $0.75 $3.00 $1.425
50,000,000 $3.75 $15.00 $7.125
100,000,000 $7.50 $30.00 $14.25
500,000,000 $37.50 $150.00 $71.25
1,000,000,000 $75.00 $300.00 $142.50
10,000,000,000 $750.00 $3,000.00 $1,425.00
100,000,000,000 $7,500.00 $30,000.00 $14,250.00

Tokens for a fixed budget

How far a fixed spend goes on GPT OSS 20B (Groq) at these list rates.

Budget Input tokens Output tokens Mixed tokens (70/30)
$0.01 133,333 33,333 70,175
$0.1 1.33M 333,333 701,754
$0.5 6.67M 1.67M 3.51M
$1.00 13.3M 3.33M 7.02M
$5.00 66.7M 16.7M 35.1M
$10.00 133.3M 33.3M 70.2M
$20.00 266.7M 66.7M 140.4M
$25.00 333.3M 83.3M 175.4M
$30.00 400M 100M 210.5M
$50.00 666.7M 166.7M 350.9M
$100.00 1.33B 333.3M 701.8M
$250.00 3.33B 833.3M 1.75B
$500.00 6.67B 1.67B 3.51B
$1,000.00 13.3B 3.33B 7.02B
$5,000.00 66.7B 16.7B 35.1B
$10,000.00 133.3B 33.3B 70.2B

At these rates

  • A workload with 70% input and 30% output costs $0.1425 per 1 million total tokens and $14.25 per 100 million.
  • At the same token count, output costs 4 times the input rate. Response length therefore deserves its own budget limit.
  • Filling the 131,072-token context window with input alone would cost about $0.00983, before any output.

Groq

Ultra-fast inference for open models with transparent per-token pricing.

All Groq models

These figures use public list prices. Cache eligibility, batch or priority modes, long-context surcharges, discounts, and regional billing can change the invoice.

Related models