Baseten
GPT OSS 120B (Baseten) token cost calculator
OpenAI gpt-oss-120b on Baseten Model APIs at low open-weight token rates.

Published list rates
Input
$0.1
per 1M tokens
Output
$0.5
per 1M tokens
Context window
128,000
tokens
Pricing source: Baseten Model APIs pricing
Token cost calculator
Estimate from text
Calculate from a token total
Project monthly volume
Uses the input/output mix above and a 30-day month: requests per day × tokens per request × 30.
Estimated cost
- Estimated tokens
- —
- Characters
- —
- Words
- —
- Input tokens
- —
- Output tokens
- —
- Input cost
- —
- Output cost
- —
- Total cost
- —
- Combined price per 1M tokens
- —
- Estimated monthly cost
- —
- Estimated daily cost
- —
Cost by token volume
Compare input-only, output-only, and 70/30 mixed workloads from one token to 100 billion. Intermediate bands make smaller production steps easier to price.
| Tokens | Input only | Output only | 70/30 mix |
|---|---|---|---|
| 1 | $0.0000001 | $0.0000005 | $0.00000022 |
| 10 | $0.000001 | $0.000005 | $0.000002 |
| 100 | $0.00001 | $0.00005 | $0.000022 |
| 200 | $0.00002 | $0.0001 | $0.000044 |
| 300 | $0.00003 | $0.00015 | $0.000066 |
| 400 | $0.00004 | $0.0002 | $0.000088 |
| 500 | $0.00005 | $0.00025 | $0.00011 |
| 1,000 | $0.0001 | $0.0005 | $0.00022 |
| 2,000 | $0.0002 | $0.001 | $0.00044 |
| 5,000 | $0.0005 | $0.0025 | $0.0011 |
| 10,000 | $0.001 | $0.005 | $0.0022 |
| 25,000 | $0.0025 | $0.0125 | $0.0055 |
| 50,000 | $0.005 | $0.025 | $0.011 |
| 100,000 | $0.01 | $0.05 | $0.022 |
| 250,000 | $0.025 | $0.125 | $0.055 |
| 500,000 | $0.05 | $0.25 | $0.11 |
| 1,000,000 | $0.1 | $0.5 | $0.22 |
| 5,000,000 | $0.5 | $2.50 | $1.10 |
| 10,000,000 | $1.00 | $5.00 | $2.20 |
| 50,000,000 | $5.00 | $25.00 | $11.00 |
| 100,000,000 | $10.00 | $50.00 | $22.00 |
| 500,000,000 | $50.00 | $250.00 | $110.00 |
| 1,000,000,000 | $100.00 | $500.00 | $220.00 |
| 10,000,000,000 | $1,000.00 | $5,000.00 | $2,200.00 |
| 100,000,000,000 | $10,000.00 | $50,000.00 | $22,000.00 |
Tokens available by budget
See how many tokens a fixed budget buys at the published rates. Display-currency values use the current exchange-rate snapshot shown on the page.
| Budget | Input tokens | Output tokens | Mixed tokens (70/30) |
|---|---|---|---|
| $0.01 | 100,000 | 20,000 | 45,455 |
| $0.1 | 1M | 200,000 | 454,545 |
| $0.5 | 5M | 1M | 2.27M |
| $1.00 | 10M | 2M | 4.55M |
| $5.00 | 50M | 10M | 22.7M |
| $10.00 | 100M | 20M | 45.5M |
| $20.00 | 200M | 40M | 90.9M |
| $25.00 | 250M | 50M | 113.6M |
| $30.00 | 300M | 60M | 136.4M |
| $50.00 | 500M | 100M | 227.3M |
| $100.00 | 1B | 200M | 454.5M |
| $250.00 | 2.5B | 500M | 1.14B |
| $500.00 | 5B | 1B | 2.27B |
| $1,000.00 | 10B | 2B | 4.55B |
| $5,000.00 | 50B | 10B | 22.7B |
| $10,000.00 | 100B | 20B | 45.5B |
Cost profile
- A workload with 70% input and 30% output costs $0.22 per 1 million total tokens and $22.00 per 100 million.
- At the same token count, output costs 5 times the input rate. Response length therefore deserves its own budget limit.
- Filling the 128,000-token context window with input alone would cost about $0.0128, before any output.
How the text estimate works
For rough planning, CostUse uses about four characters per token and about 0.75 words per token. Actual tokenization changes with language, model, formatting, and special tokens.
These results use public list prices. Cache eligibility, batch or priority modes, long-context surcharges, discounts, and regional billing can change the final invoice.
Questions about this rate
How much is 1 million mixed tokens on GPT OSS 120B (Baseten)?
About $0.22 at a 70% input / 30% output mix using list prices verified 2026-07-30.
How much are 100 million mixed tokens?
About $22.00 at the same 70/30 mix — useful for rough monthly volume planning.
What is the GPT OSS 120B (Baseten) price per million input and output tokens?
Input $0.1 and output $0.5 per 1M tokens (USD list rates). Use the calculator above with your real token mix.
Is the text-to-token estimate exact?
No. The estimator approximates tokens from characters and words. Real bills depend on the provider tokenizer, special tokens, and discounts (cache, batch, priority).
What does cost per token mean, and why is output usually more expensive?
Providers charge for tokens processed. Output tokens (generated replies) often cost several times more than input tokens (your prompt), so completion length can dominate the bill.