SiliconFlow

Kimi K3 (SiliconFlow) token cost

Kimi K3 on SiliconFlow — multi-host compare with Baseten and Moonshot rates.

Published list rates

Input

$3.00

per 1M tokens

Output

$15.00

per 1M tokens

Cached input

$0.3

per 1M tokens

Context window

1,049,000

tokens

Pricing source: SiliconFlow pricing

Estimate cost

Estimate from text

Calculate from a token total

Project monthly volume

Daily requests × tokens per request × 30 days.

Estimated cost

Estimated tokens
Characters
Words
Input tokens
Output tokens
Input cost
Output cost
Total cost
Combined price per 1M tokens
Estimated monthly cost
Estimated daily cost

Cost as volume grows

Kimi K3 (SiliconFlow) at rising volume: prompt-only, reply-only, and a mixed chat (70% in, 30% out).

Tokens Input only Output only Mix (70% in / 30% out)
1 $0.000003 $0.000015 $0.000007
10 $0.00003 $0.00015 $0.000066
100 $0.0003 $0.0015 $0.00066
200 $0.0006 $0.003 $0.00132
300 $0.0009 $0.0045 $0.00198
400 $0.0012 $0.006 $0.00264
500 $0.0015 $0.0075 $0.0033
1,000 $0.003 $0.015 $0.0066
2,000 $0.006 $0.03 $0.0132
5,000 $0.015 $0.075 $0.033
10,000 $0.03 $0.15 $0.066
25,000 $0.075 $0.375 $0.165
50,000 $0.15 $0.75 $0.33
100,000 $0.3 $1.50 $0.66
250,000 $0.75 $3.75 $1.65
500,000 $1.50 $7.50 $3.30
1,000,000 $3.00 $15.00 $6.60
5,000,000 $15.00 $75.00 $33.00
10,000,000 $30.00 $150.00 $66.00
50,000,000 $150.00 $750.00 $330.00
100,000,000 $300.00 $1,500.00 $660.00
500,000,000 $1,500.00 $7,500.00 $3,300.00
1,000,000,000 $3,000.00 $15,000.00 $6,600.00
10,000,000,000 $30,000.00 $150,000.00 $66,000.00
100,000,000,000 $300,000.00 $1,500,000.00 $660,000.00

Tokens for a fixed budget

If you cap spend, this is roughly how many Kimi K3 (SiliconFlow) tokens that money buys.

Budget Input tokens Output tokens Mixed tokens (70/30)
$0.01 3,333 667 1,515
$0.1 33,333 6,667 15,152
$0.5 166,667 33,333 75,758
$1.00 333,333 66,667 151,515
$5.00 1.67M 333,333 757,576
$10.00 3.33M 666,667 1.52M
$20.00 6.67M 1.33M 3.03M
$25.00 8.33M 1.67M 3.79M
$30.00 10M 2M 4.55M
$50.00 16.7M 3.33M 7.58M
$100.00 33.3M 6.67M 15.2M
$250.00 83.3M 16.7M 37.9M
$500.00 166.7M 33.3M 75.8M
$1,000.00 333.3M 66.7M 151.5M
$5,000.00 1.67B 333.3M 757.6M
$10,000.00 3.33B 666.7M 1.52B

At these rates

  • A workload with 70% input and 30% output costs $6.60 per 1 million total tokens and $660.00 per 100 million.
  • At the same token count, output costs 5 times the input rate. Response length therefore deserves its own budget limit.
  • Filling the 1,049,000-token context window with input alone would cost about $3.147, before any output.
  • Cached input is $0.3 per million, 90% below standard input when the workload meets the provider’s cache rules.

SiliconFlow

Serverless open-model inference with public per-token rates for DeepSeek, Qwen, GLM, and more.

All SiliconFlow models

These figures use public list prices. Cache eligibility, batch or priority modes, long-context surcharges, discounts, and regional billing can change the invoice.

Related models