SiliconFlow

GLM-5.2 (SiliconFlow) token cost

GLM-5.2 on SiliconFlow for long-context coding agents with cache reads.

Published list rates

Input

$1.302

per 1M tokens

Output

$4.092

per 1M tokens

Cached input

$0.26

per 1M tokens

Context window

1,049,000

tokens

Pricing source: SiliconFlow pricing

Estimate cost

Estimate from text

Calculate from a token total

Project monthly volume

Daily requests × tokens per request × 30 days.

Estimated cost

Estimated tokens
Characters
Words
Input tokens
Output tokens
Input cost
Output cost
Total cost
Combined price per 1M tokens
Estimated monthly cost
Estimated daily cost

Cost as volume grows

GLM-5.2 (SiliconFlow) at rising volume: prompt-only, reply-only, and a mixed chat (70% in, 30% out).

Tokens Input only Output only Mix (70% in / 30% out)
1 $0.000001 $0.000004 $0.000002
10 $0.000013 $0.000041 $0.000021
100 $0.00013 $0.000409 $0.000214
200 $0.00026 $0.000818 $0.000428
300 $0.000391 $0.001228 $0.000642
400 $0.000521 $0.001637 $0.000856
500 $0.000651 $0.002046 $0.00107
1,000 $0.001302 $0.004092 $0.002139
2,000 $0.002604 $0.008184 $0.004278
5,000 $0.00651 $0.02046 $0.0107
10,000 $0.01302 $0.04092 $0.02139
25,000 $0.03255 $0.1023 $0.05347
50,000 $0.0651 $0.2046 $0.10695
100,000 $0.1302 $0.4092 $0.2139
250,000 $0.3255 $1.023 $0.53475
500,000 $0.651 $2.046 $1.0695
1,000,000 $1.302 $4.092 $2.139
5,000,000 $6.51 $20.46 $10.695
10,000,000 $13.02 $40.92 $21.39
50,000,000 $65.10 $204.60 $106.95
100,000,000 $130.20 $409.20 $213.90
500,000,000 $651.00 $2,046.00 $1,069.50
1,000,000,000 $1,302.00 $4,092.00 $2,139.00
10,000,000,000 $13,020.00 $40,920.00 $21,390.00
100,000,000,000 $130,200.00 $409,200.00 $213,900.00

Tokens for a fixed budget

How far a fixed spend goes on GLM-5.2 (SiliconFlow) at these list rates.

Budget Input tokens Output tokens Mixed tokens (70/30)
$0.01 7,680 2,444 4,675
$0.1 76,805 24,438 46,751
$0.5 384,025 122,190 233,754
$1.00 768,049 244,379 467,508
$5.00 3.84M 1.22M 2.34M
$10.00 7.68M 2.44M 4.68M
$20.00 15.4M 4.89M 9.35M
$25.00 19.2M 6.11M 11.7M
$30.00 23M 7.33M 14M
$50.00 38.4M 12.2M 23.4M
$100.00 76.8M 24.4M 46.8M
$250.00 192M 61.1M 116.9M
$500.00 384M 122.2M 233.8M
$1,000.00 768M 244.4M 467.5M
$5,000.00 3.84B 1.22B 2.34B
$10,000.00 7.68B 2.44B 4.68B

At these rates

  • A workload with 70% input and 30% output costs $2.139 per 1 million total tokens and $213.90 per 100 million.
  • At the same token count, output costs 3.14 times the input rate. Response length therefore deserves its own budget limit.
  • Filling the 1,049,000-token context window with input alone would cost about $1.3658, before any output.
  • Cached input is $0.26 per million, 80% below standard input when the workload meets the provider’s cache rules.

SiliconFlow

Serverless open-model inference with public per-token rates for DeepSeek, Qwen, GLM, and more.

All SiliconFlow models

These figures use public list prices. Cache eligibility, batch or priority modes, long-context surcharges, discounts, and regional billing can change the invoice.

Related models