SiliconFlow

gpt-oss-120b (SiliconFlow) token cost

OpenAI gpt-oss-120b on SiliconFlow at competitive open-weight serverless rates.

Published list rates

Input

$0.05

per 1M tokens

Output

$0.45

per 1M tokens

Context window

131,000

tokens

Pricing source: SiliconFlow pricing

Estimate cost

Estimate from text

Calculate from a token total

Project monthly volume

Daily requests × tokens per request × 30 days.

Estimated cost

Estimated tokens
Characters
Words
Input tokens
Output tokens
Input cost
Output cost
Total cost
Combined price per 1M tokens
Estimated monthly cost
Estimated daily cost

Cost as volume grows

What gpt-oss-120b (SiliconFlow) costs if you only send tokens, only get a reply, or do both (70% in / 30% out). Rows start small and go up to huge traffic.

Tokens Input only Output only Mix (70% in / 30% out)
1 $0.00000005 $0.00000045 $0.00000017
10 $0.0000005 $0.000005 $0.000002
100 $0.000005 $0.000045 $0.000017
200 $0.00001 $0.00009 $0.000034
300 $0.000015 $0.000135 $0.000051
400 $0.00002 $0.00018 $0.000068
500 $0.000025 $0.000225 $0.000085
1,000 $0.00005 $0.00045 $0.00017
2,000 $0.0001 $0.0009 $0.00034
5,000 $0.00025 $0.00225 $0.00085
10,000 $0.0005 $0.0045 $0.0017
25,000 $0.00125 $0.01125 $0.00425
50,000 $0.0025 $0.0225 $0.0085
100,000 $0.005 $0.045 $0.017
250,000 $0.0125 $0.1125 $0.0425
500,000 $0.025 $0.225 $0.085
1,000,000 $0.05 $0.45 $0.17
5,000,000 $0.25 $2.25 $0.85
10,000,000 $0.5 $4.50 $1.70
50,000,000 $2.50 $22.50 $8.50
100,000,000 $5.00 $45.00 $17.00
500,000,000 $25.00 $225.00 $85.00
1,000,000,000 $50.00 $450.00 $170.00
10,000,000,000 $500.00 $4,500.00 $1,700.00
100,000,000,000 $5,000.00 $45,000.00 $17,000.00

Tokens for a fixed budget

If you cap spend, this is roughly how many gpt-oss-120b (SiliconFlow) tokens that money buys.

Budget Input tokens Output tokens Mixed tokens (70/30)
$0.01 200,000 22,222 58,824
$0.1 2M 222,222 588,235
$0.5 10M 1.11M 2.94M
$1.00 20M 2.22M 5.88M
$5.00 100M 11.1M 29.4M
$10.00 200M 22.2M 58.8M
$20.00 400M 44.4M 117.6M
$25.00 500M 55.6M 147.1M
$30.00 600M 66.7M 176.5M
$50.00 1B 111.1M 294.1M
$100.00 2B 222.2M 588.2M
$250.00 5B 555.6M 1.47B
$500.00 10B 1.11B 2.94B
$1,000.00 20B 2.22B 5.88B
$5,000.00 100B 11.1B 29.4B
$10,000.00 200B 22.2B 58.8B

At these rates

  • A workload with 70% input and 30% output costs $0.17 per 1 million total tokens and $17.00 per 100 million.
  • At the same token count, output costs 9 times the input rate. Response length therefore deserves its own budget limit.
  • Filling the 131,000-token context window with input alone would cost about $0.00655, before any output.

SiliconFlow

Serverless open-model inference with public per-token rates for DeepSeek, Qwen, GLM, and more.

All SiliconFlow models

These figures use public list prices. Cache eligibility, batch or priority modes, long-context surcharges, discounts, and regional billing can change the invoice.

Related models