SiliconFlow
GLM-5.2 (SiliconFlow) token cost
GLM-5.2 on SiliconFlow for long-context coding agents with cache reads.

Published list rates
Input
$1.302
per 1M tokens
Output
$4.092
per 1M tokens
Cached input
$0.26
per 1M tokens
Context window
1,049,000
tokens
Pricing source: SiliconFlow pricing
Estimate cost
Estimate from text
Calculate from a token total
Project monthly volume
Daily requests × tokens per request × 30 days.
Estimated cost
- Estimated tokens
- —
- Characters
- —
- Words
- —
- Input tokens
- —
- Output tokens
- —
- Input cost
- —
- Output cost
- —
- Total cost
- —
- Combined price per 1M tokens
- —
- Estimated monthly cost
- —
- Estimated daily cost
- —
Cost as volume grows
GLM-5.2 (SiliconFlow) at rising volume: prompt-only, reply-only, and a mixed chat (70% in, 30% out).
| Tokens | Input only | Output only | Mix (70% in / 30% out) |
|---|---|---|---|
| 1 | $0.000001 | $0.000004 | $0.000002 |
| 10 | $0.000013 | $0.000041 | $0.000021 |
| 100 | $0.00013 | $0.000409 | $0.000214 |
| 200 | $0.00026 | $0.000818 | $0.000428 |
| 300 | $0.000391 | $0.001228 | $0.000642 |
| 400 | $0.000521 | $0.001637 | $0.000856 |
| 500 | $0.000651 | $0.002046 | $0.00107 |
| 1,000 | $0.001302 | $0.004092 | $0.002139 |
| 2,000 | $0.002604 | $0.008184 | $0.004278 |
| 5,000 | $0.00651 | $0.02046 | $0.0107 |
| 10,000 | $0.01302 | $0.04092 | $0.02139 |
| 25,000 | $0.03255 | $0.1023 | $0.05347 |
| 50,000 | $0.0651 | $0.2046 | $0.10695 |
| 100,000 | $0.1302 | $0.4092 | $0.2139 |
| 250,000 | $0.3255 | $1.023 | $0.53475 |
| 500,000 | $0.651 | $2.046 | $1.0695 |
| 1,000,000 | $1.302 | $4.092 | $2.139 |
| 5,000,000 | $6.51 | $20.46 | $10.695 |
| 10,000,000 | $13.02 | $40.92 | $21.39 |
| 50,000,000 | $65.10 | $204.60 | $106.95 |
| 100,000,000 | $130.20 | $409.20 | $213.90 |
| 500,000,000 | $651.00 | $2,046.00 | $1,069.50 |
| 1,000,000,000 | $1,302.00 | $4,092.00 | $2,139.00 |
| 10,000,000,000 | $13,020.00 | $40,920.00 | $21,390.00 |
| 100,000,000,000 | $130,200.00 | $409,200.00 | $213,900.00 |
Tokens for a fixed budget
How far a fixed spend goes on GLM-5.2 (SiliconFlow) at these list rates.
| Budget | Input tokens | Output tokens | Mixed tokens (70/30) |
|---|---|---|---|
| $0.01 | 7,680 | 2,444 | 4,675 |
| $0.1 | 76,805 | 24,438 | 46,751 |
| $0.5 | 384,025 | 122,190 | 233,754 |
| $1.00 | 768,049 | 244,379 | 467,508 |
| $5.00 | 3.84M | 1.22M | 2.34M |
| $10.00 | 7.68M | 2.44M | 4.68M |
| $20.00 | 15.4M | 4.89M | 9.35M |
| $25.00 | 19.2M | 6.11M | 11.7M |
| $30.00 | 23M | 7.33M | 14M |
| $50.00 | 38.4M | 12.2M | 23.4M |
| $100.00 | 76.8M | 24.4M | 46.8M |
| $250.00 | 192M | 61.1M | 116.9M |
| $500.00 | 384M | 122.2M | 233.8M |
| $1,000.00 | 768M | 244.4M | 467.5M |
| $5,000.00 | 3.84B | 1.22B | 2.34B |
| $10,000.00 | 7.68B | 2.44B | 4.68B |
At these rates
- A workload with 70% input and 30% output costs $2.139 per 1 million total tokens and $213.90 per 100 million.
- At the same token count, output costs 3.14 times the input rate. Response length therefore deserves its own budget limit.
- Filling the 1,049,000-token context window with input alone would cost about $1.3658, before any output.
- Cached input is $0.26 per million, 80% below standard input when the workload meets the provider’s cache rules.
These figures use public list prices. Cache eligibility, batch or priority modes, long-context surcharges, discounts, and regional billing can change the invoice.