OpenAI
GPT-5.4 nano token cost
Smallest GPT-5.4 class model for latency-sensitive bulk classification and routing.

Published list rates
Input
$0.2
per 1M tokens
Output
$1.25
per 1M tokens
Cached input
$0.02
per 1M tokens
Context window
400,000
tokens
Pricing source: OpenAI API pricing
Estimate cost
Estimate from text
Calculate from a token total
Project monthly volume
Daily requests × tokens per request × 30 days.
Estimated cost
- Estimated tokens
- —
- Characters
- —
- Words
- —
- Input tokens
- —
- Output tokens
- —
- Input cost
- —
- Output cost
- —
- Total cost
- —
- Combined price per 1M tokens
- —
- Estimated monthly cost
- —
- Estimated daily cost
- —
Cost as volume grows
Read down to see GPT-5.4 nano get more expensive as tokens grow. The last column is a typical mix: 70% input, 30% output.
| Tokens | Input only | Output only | Mix (70% in / 30% out) |
|---|---|---|---|
| 1 | $0.0000002 | $0.000001 | $0.000000515 |
| 10 | $0.000002 | $0.000013 | $0.000005 |
| 100 | $0.00002 | $0.000125 | $0.000052 |
| 200 | $0.00004 | $0.00025 | $0.000103 |
| 300 | $0.00006 | $0.000375 | $0.000155 |
| 400 | $0.00008 | $0.0005 | $0.000206 |
| 500 | $0.0001 | $0.000625 | $0.000257 |
| 1,000 | $0.0002 | $0.00125 | $0.000515 |
| 2,000 | $0.0004 | $0.0025 | $0.00103 |
| 5,000 | $0.001 | $0.00625 | $0.002575 |
| 10,000 | $0.002 | $0.0125 | $0.00515 |
| 25,000 | $0.005 | $0.03125 | $0.01288 |
| 50,000 | $0.01 | $0.0625 | $0.02575 |
| 100,000 | $0.02 | $0.125 | $0.0515 |
| 250,000 | $0.05 | $0.3125 | $0.12875 |
| 500,000 | $0.1 | $0.625 | $0.2575 |
| 1,000,000 | $0.2 | $1.25 | $0.515 |
| 5,000,000 | $1.00 | $6.25 | $2.575 |
| 10,000,000 | $2.00 | $12.50 | $5.15 |
| 50,000,000 | $10.00 | $62.50 | $25.75 |
| 100,000,000 | $20.00 | $125.00 | $51.50 |
| 500,000,000 | $100.00 | $625.00 | $257.50 |
| 1,000,000,000 | $200.00 | $1,250.00 | $515.00 |
| 10,000,000,000 | $2,000.00 | $12,500.00 | $5,150.00 |
| 100,000,000,000 | $20,000.00 | $125,000.00 | $51,500.00 |
Tokens for a fixed budget
How far a fixed spend goes on GPT-5.4 nano at these list rates.
| Budget | Input tokens | Output tokens | Mixed tokens (70/30) |
|---|---|---|---|
| $0.01 | 50,000 | 8,000 | 19,417 |
| $0.1 | 500,000 | 80,000 | 194,175 |
| $0.5 | 2.5M | 400,000 | 970,874 |
| $1.00 | 5M | 800,000 | 1.94M |
| $5.00 | 25M | 4M | 9.71M |
| $10.00 | 50M | 8M | 19.4M |
| $20.00 | 100M | 16M | 38.8M |
| $25.00 | 125M | 20M | 48.5M |
| $30.00 | 150M | 24M | 58.3M |
| $50.00 | 250M | 40M | 97.1M |
| $100.00 | 500M | 80M | 194.2M |
| $250.00 | 1.25B | 200M | 485.4M |
| $500.00 | 2.5B | 400M | 970.9M |
| $1,000.00 | 5B | 800M | 1.94B |
| $5,000.00 | 25B | 4B | 9.71B |
| $10,000.00 | 50B | 8B | 19.4B |
At these rates
- A workload with 70% input and 30% output costs $0.515 per 1 million total tokens and $51.50 per 100 million.
- At the same token count, output costs 6.25 times the input rate. Response length therefore deserves its own budget limit.
- Filling the 400,000-token context window with input alone would cost about $0.08, before any output.
- Cached input is $0.02 per million, 90% below standard input when the workload meets the provider’s cache rules.
These figures use public list prices. Cache eligibility, batch or priority modes, long-context surcharges, discounts, and regional billing can change the invoice.