SiliconFlow

DeepSeek V3.2 (SiliconFlow) token cost

DeepSeek V3.2 on SiliconFlow with cache discounts for production open routing.

Published list rates

Input

$0.27

per 1M tokens

Output

$0.42

per 1M tokens

Cached input

$0.135

per 1M tokens

Context window

164,000

tokens

Pricing source: SiliconFlow pricing

Estimate cost

Estimate from text

Calculate from a token total

Project monthly volume

Daily requests × tokens per request × 30 days.

Estimated cost

Estimated tokens
Characters
Words
Input tokens
Output tokens
Input cost
Output cost
Total cost
Combined price per 1M tokens
Estimated monthly cost
Estimated daily cost

Cost as volume grows

What DeepSeek V3.2 (SiliconFlow) costs if you only send tokens, only get a reply, or do both (70% in / 30% out). Rows start small and go up to huge traffic.

Tokens Input only Output only Mix (70% in / 30% out)
1 $0.00000027 $0.00000042 $0.000000315
10 $0.000003 $0.000004 $0.000003
100 $0.000027 $0.000042 $0.000032
200 $0.000054 $0.000084 $0.000063
300 $0.000081 $0.000126 $0.000095
400 $0.000108 $0.000168 $0.000126
500 $0.000135 $0.00021 $0.000157
1,000 $0.00027 $0.00042 $0.000315
2,000 $0.00054 $0.00084 $0.00063
5,000 $0.00135 $0.0021 $0.001575
10,000 $0.0027 $0.0042 $0.00315
25,000 $0.00675 $0.0105 $0.007875
50,000 $0.0135 $0.021 $0.01575
100,000 $0.027 $0.042 $0.0315
250,000 $0.0675 $0.105 $0.07875
500,000 $0.135 $0.21 $0.1575
1,000,000 $0.27 $0.42 $0.315
5,000,000 $1.35 $2.10 $1.575
10,000,000 $2.70 $4.20 $3.15
50,000,000 $13.50 $21.00 $15.75
100,000,000 $27.00 $42.00 $31.50
500,000,000 $135.00 $210.00 $157.50
1,000,000,000 $270.00 $420.00 $315.00
10,000,000,000 $2,700.00 $4,200.00 $3,150.00
100,000,000,000 $27,000.00 $42,000.00 $31,500.00

Tokens for a fixed budget

If you cap spend, this is roughly how many DeepSeek V3.2 (SiliconFlow) tokens that money buys.

Budget Input tokens Output tokens Mixed tokens (70/30)
$0.01 37,037 23,810 31,746
$0.1 370,370 238,095 317,460
$0.5 1.85M 1.19M 1.59M
$1.00 3.7M 2.38M 3.17M
$5.00 18.5M 11.9M 15.9M
$10.00 37M 23.8M 31.7M
$20.00 74.1M 47.6M 63.5M
$25.00 92.6M 59.5M 79.4M
$30.00 111.1M 71.4M 95.2M
$50.00 185.2M 119M 158.7M
$100.00 370.4M 238.1M 317.5M
$250.00 925.9M 595.2M 793.7M
$500.00 1.85B 1.19B 1.59B
$1,000.00 3.7B 2.38B 3.17B
$5,000.00 18.5B 11.9B 15.9B
$10,000.00 37B 23.8B 31.7B

At these rates

  • A workload with 70% input and 30% output costs $0.315 per 1 million total tokens and $31.50 per 100 million.
  • At the same token count, output costs 1.56 times the input rate. Response length therefore deserves its own budget limit.
  • Filling the 164,000-token context window with input alone would cost about $0.04428, before any output.
  • Cached input is $0.135 per million, 50% below standard input when the workload meets the provider’s cache rules.

SiliconFlow

Serverless open-model inference with public per-token rates for DeepSeek, Qwen, GLM, and more.

All SiliconFlow models

These figures use public list prices. Cache eligibility, batch or priority modes, long-context surcharges, discounts, and regional billing can change the invoice.

Related models