Inception

Mercury 2 token cost

Inception Mercury 2 diffusion LLM for high-throughput chat with cached input discounts.

Published list rates

Input

$0.25

per 1M tokens

Output

$0.75

per 1M tokens

Cached input

$0.025

per 1M tokens

Context window

128,000

tokens

Pricing source: Inception models & pricing

Estimate cost

Estimate from text

Calculate from a token total

Project monthly volume

Daily requests × tokens per request × 30 days.

Estimated cost

Estimated tokens
Characters
Words
Input tokens
Output tokens
Input cost
Output cost
Total cost
Combined price per 1M tokens
Estimated monthly cost
Estimated daily cost

Cost as volume grows

Read down to see Mercury 2 get more expensive as tokens grow. The last column is a typical mix: 70% input, 30% output.

Tokens Input only Output only Mix (70% in / 30% out)
1 $0.00000025 $0.00000075 $0.0000004
10 $0.000003 $0.000008 $0.000004
100 $0.000025 $0.000075 $0.00004
200 $0.00005 $0.00015 $0.00008
300 $0.000075 $0.000225 $0.00012
400 $0.0001 $0.0003 $0.00016
500 $0.000125 $0.000375 $0.0002
1,000 $0.00025 $0.00075 $0.0004
2,000 $0.0005 $0.0015 $0.0008
5,000 $0.00125 $0.00375 $0.002
10,000 $0.0025 $0.0075 $0.004
25,000 $0.00625 $0.01875 $0.01
50,000 $0.0125 $0.0375 $0.02
100,000 $0.025 $0.075 $0.04
250,000 $0.0625 $0.1875 $0.1
500,000 $0.125 $0.375 $0.2
1,000,000 $0.25 $0.75 $0.4
5,000,000 $1.25 $3.75 $2.00
10,000,000 $2.50 $7.50 $4.00
50,000,000 $12.50 $37.50 $20.00
100,000,000 $25.00 $75.00 $40.00
500,000,000 $125.00 $375.00 $200.00
1,000,000,000 $250.00 $750.00 $400.00
10,000,000,000 $2,500.00 $7,500.00 $4,000.00
100,000,000,000 $25,000.00 $75,000.00 $40,000.00

Tokens for a fixed budget

How far a fixed spend goes on Mercury 2 at these list rates.

Budget Input tokens Output tokens Mixed tokens (70/30)
$0.01 40,000 13,333 25,000
$0.1 400,000 133,333 250,000
$0.5 2M 666,667 1.25M
$1.00 4M 1.33M 2.5M
$5.00 20M 6.67M 12.5M
$10.00 40M 13.3M 25M
$20.00 80M 26.7M 50M
$25.00 100M 33.3M 62.5M
$30.00 120M 40M 75M
$50.00 200M 66.7M 125M
$100.00 400M 133.3M 250M
$250.00 1B 333.3M 625M
$500.00 2B 666.7M 1.25B
$1,000.00 4B 1.33B 2.5B
$5,000.00 20B 6.67B 12.5B
$10,000.00 40B 13.3B 25B

At these rates

  • A workload with 70% input and 30% output costs $0.4 per 1 million total tokens and $40.00 per 100 million.
  • At the same token count, output costs 3 times the input rate. Response length therefore deserves its own budget limit.
  • Filling the 128,000-token context window with input alone would cost about $0.032, before any output.
  • Cached input is $0.025 per million, 90% below standard input when the workload meets the provider’s cache rules.

Inception

Diffusion LLMs (Mercury) with public per-token API rates for high-throughput generation.

All Inception models

New accounts include free trial tokens. Confirm live Inception docs before production budgets.

These figures use public list prices. Cache eligibility, batch or priority modes, long-context surcharges, discounts, and regional billing can change the invoice.

Related models