Qwen

Qwen3.8 Max token cost

Alibaba’s current Qwen flagship for long-horizon coding and professional work, priced below Qwen3.7 Max on the international sheet.

Published list rates

Input

$2.00

per 1M tokens

Output

$6.00

per 1M tokens

Cached input

$0.25

per 1M tokens

Context window

1,000,000

tokens

Pricing source: QwenCloud Qwen3.8-Max pricing

Estimate cost

Estimate from text

Calculate from a token total

Project monthly volume

Daily requests × tokens per request × 30 days.

Estimated cost

Estimated tokens
Characters
Words
Input tokens
Output tokens
Input cost
Output cost
Total cost
Combined price per 1M tokens
Estimated monthly cost
Estimated daily cost

Cost as volume grows

Qwen3.8 Max at rising volume: prompt-only, reply-only, and a mixed chat (70% in, 30% out).

Tokens Input only Output only Mix (70% in / 30% out)
1 $0.000002 $0.000006 $0.000003
10 $0.00002 $0.00006 $0.000032
100 $0.0002 $0.0006 $0.00032
200 $0.0004 $0.0012 $0.00064
300 $0.0006 $0.0018 $0.00096
400 $0.0008 $0.0024 $0.00128
500 $0.001 $0.003 $0.0016
1,000 $0.002 $0.006 $0.0032
2,000 $0.004 $0.012 $0.0064
5,000 $0.01 $0.03 $0.016
10,000 $0.02 $0.06 $0.032
25,000 $0.05 $0.15 $0.08
50,000 $0.1 $0.3 $0.16
100,000 $0.2 $0.6 $0.32
250,000 $0.5 $1.50 $0.8
500,000 $1.00 $3.00 $1.60
1,000,000 $2.00 $6.00 $3.20
5,000,000 $10.00 $30.00 $16.00
10,000,000 $20.00 $60.00 $32.00
50,000,000 $100.00 $300.00 $160.00
100,000,000 $200.00 $600.00 $320.00
500,000,000 $1,000.00 $3,000.00 $1,600.00
1,000,000,000 $2,000.00 $6,000.00 $3,200.00
10,000,000,000 $20,000.00 $60,000.00 $32,000.00
100,000,000,000 $200,000.00 $600,000.00 $320,000.00

Tokens for a fixed budget

If you cap spend, this is roughly how many Qwen3.8 Max tokens that money buys.

Budget Input tokens Output tokens Mixed tokens (70/30)
$0.01 5,000 1,667 3,125
$0.1 50,000 16,667 31,250
$0.5 250,000 83,333 156,250
$1.00 500,000 166,667 312,500
$5.00 2.5M 833,333 1.56M
$10.00 5M 1.67M 3.13M
$20.00 10M 3.33M 6.25M
$25.00 12.5M 4.17M 7.81M
$30.00 15M 5M 9.38M
$50.00 25M 8.33M 15.6M
$100.00 50M 16.7M 31.3M
$250.00 125M 41.7M 78.1M
$500.00 250M 83.3M 156.3M
$1,000.00 500M 166.7M 312.5M
$5,000.00 2.5B 833.3M 1.56B
$10,000.00 5B 1.67B 3.13B

At these rates

  • A workload with 70% input and 30% output costs $3.20 per 1 million total tokens and $320.00 per 100 million.
  • At the same token count, output costs 3 times the input rate. Response length therefore deserves its own budget limit.
  • Filling the 1,000,000-token context window with input alone would cost about $2.00, before any output.
  • Cached input is $0.25 per million, 88% below standard input when the workload meets the provider’s cache rules.

Qwen

Alibaba Cloud Model Studio first-party Qwen models with international token rates.

All Qwen models

Implicit cache reads are $0.25 per 1M. Explicit cache creation is $2.50 per 1M and explicit cache reads are $0.17 per 1M. Context is 1M with about 131k max output.

These figures use public list prices. Cache eligibility, batch or priority modes, long-context surcharges, discounts, and regional billing can change the invoice.

Related models