IBM
Mistral Large (watsonx.ai) token cost
Mistral Large 2512 on watsonx.ai for stronger reasoning and multilingual enterprise work.

Published list rates
Input
$0.636
per 1M tokens
Output
$1.908
per 1M tokens
Context window
128,000
tokens
Pricing source: IBM watsonx.ai pricing
Estimate cost
Estimate from text
Calculate from a token total
Project monthly volume
Daily requests × tokens per request × 30 days.
Estimated cost
- Estimated tokens
- —
- Characters
- —
- Words
- —
- Input tokens
- —
- Output tokens
- —
- Input cost
- —
- Output cost
- —
- Total cost
- —
- Combined price per 1M tokens
- —
- Estimated monthly cost
- —
- Estimated daily cost
- —
Cost as volume grows
Mistral Large (watsonx.ai) at rising volume: prompt-only, reply-only, and a mixed chat (70% in, 30% out).
| Tokens | Input only | Output only | Mix (70% in / 30% out) |
|---|---|---|---|
| 1 | $0.000000636 | $0.000002 | $0.000001 |
| 10 | $0.000006 | $0.000019 | $0.00001 |
| 100 | $0.000064 | $0.000191 | $0.000102 |
| 200 | $0.000127 | $0.000382 | $0.000204 |
| 300 | $0.000191 | $0.000572 | $0.000305 |
| 400 | $0.000254 | $0.000763 | $0.000407 |
| 500 | $0.000318 | $0.000954 | $0.000509 |
| 1,000 | $0.000636 | $0.001908 | $0.001018 |
| 2,000 | $0.001272 | $0.003816 | $0.002035 |
| 5,000 | $0.00318 | $0.00954 | $0.005088 |
| 10,000 | $0.00636 | $0.01908 | $0.01018 |
| 25,000 | $0.0159 | $0.0477 | $0.02544 |
| 50,000 | $0.0318 | $0.0954 | $0.05088 |
| 100,000 | $0.0636 | $0.1908 | $0.10176 |
| 250,000 | $0.159 | $0.477 | $0.2544 |
| 500,000 | $0.318 | $0.954 | $0.5088 |
| 1,000,000 | $0.636 | $1.908 | $1.0176 |
| 5,000,000 | $3.18 | $9.54 | $5.088 |
| 10,000,000 | $6.36 | $19.08 | $10.176 |
| 50,000,000 | $31.80 | $95.40 | $50.88 |
| 100,000,000 | $63.60 | $190.80 | $101.76 |
| 500,000,000 | $318.00 | $954.00 | $508.80 |
| 1,000,000,000 | $636.00 | $1,908.00 | $1,017.60 |
| 10,000,000,000 | $6,360.00 | $19,080.00 | $10,176.00 |
| 100,000,000,000 | $63,600.00 | $190,800.00 | $101,760.00 |
Tokens for a fixed budget
Tokens Mistral Large (watsonx.ai) can process for a set budget — send-only, reply-only, and a 70/30 mix.
| Budget | Input tokens | Output tokens | Mixed tokens (70/30) |
|---|---|---|---|
| $0.01 | 15,723 | 5,241 | 9,827 |
| $0.1 | 157,233 | 52,411 | 98,270 |
| $0.5 | 786,164 | 262,055 | 491,352 |
| $1.00 | 1.57M | 524,109 | 982,704 |
| $5.00 | 7.86M | 2.62M | 4.91M |
| $10.00 | 15.7M | 5.24M | 9.83M |
| $20.00 | 31.4M | 10.5M | 19.7M |
| $25.00 | 39.3M | 13.1M | 24.6M |
| $30.00 | 47.2M | 15.7M | 29.5M |
| $50.00 | 78.6M | 26.2M | 49.1M |
| $100.00 | 157.2M | 52.4M | 98.3M |
| $250.00 | 393.1M | 131M | 245.7M |
| $500.00 | 786.2M | 262.1M | 491.4M |
| $1,000.00 | 1.57B | 524.1M | 982.7M |
| $5,000.00 | 7.86B | 2.62B | 4.91B |
| $10,000.00 | 15.7B | 5.24B | 9.83B |
At these rates
- A workload with 70% input and 30% output costs $1.0176 per 1 million total tokens and $101.76 per 100 million.
- At the same token count, output costs 3 times the input rate. Response length therefore deserves its own budget limit.
- Filling the 128,000-token context window with input alone would cost about $0.08141, before any output.
These figures use public list prices. Cache eligibility, batch or priority modes, long-context surcharges, discounts, and regional billing can change the invoice.