OpenAI

GPT-4.1 mini token cost

Long-context GPT-4.1 mini still listed for high-volume light workloads.

Published list rates

Input

$0.4

per 1M tokens

Output

$1.60

per 1M tokens

Cached input

$0.04

per 1M tokens

Context window

1,000,000

tokens

Pricing source: OpenAI API pricing

Estimate cost

Estimate from text

Calculate from a token total

Project monthly volume

Daily requests × tokens per request × 30 days.

Estimated cost

Estimated tokens
Characters
Words
Input tokens
Output tokens
Input cost
Output cost
Total cost
Combined price per 1M tokens
Estimated monthly cost
Estimated daily cost

Cost as volume grows

Three views of GPT-4.1 mini: sending only, receiving only, and a 70/30 mix. Same list rates, bigger rows.

Tokens Input only Output only Mix (70% in / 30% out)
1 $0.0000004 $0.000002 $0.00000076
10 $0.000004 $0.000016 $0.000008
100 $0.00004 $0.00016 $0.000076
200 $0.00008 $0.00032 $0.000152
300 $0.00012 $0.00048 $0.000228
400 $0.00016 $0.00064 $0.000304
500 $0.0002 $0.0008 $0.00038
1,000 $0.0004 $0.0016 $0.00076
2,000 $0.0008 $0.0032 $0.00152
5,000 $0.002 $0.008 $0.0038
10,000 $0.004 $0.016 $0.0076
25,000 $0.01 $0.04 $0.019
50,000 $0.02 $0.08 $0.038
100,000 $0.04 $0.16 $0.076
250,000 $0.1 $0.4 $0.19
500,000 $0.2 $0.8 $0.38
1,000,000 $0.4 $1.60 $0.76
5,000,000 $2.00 $8.00 $3.80
10,000,000 $4.00 $16.00 $7.60
50,000,000 $20.00 $80.00 $38.00
100,000,000 $40.00 $160.00 $76.00
500,000,000 $200.00 $800.00 $380.00
1,000,000,000 $400.00 $1,600.00 $760.00
10,000,000,000 $4,000.00 $16,000.00 $7,600.00
100,000,000,000 $40,000.00 $160,000.00 $76,000.00

Tokens for a fixed budget

Tokens GPT-4.1 mini can process for a set budget — send-only, reply-only, and a 70/30 mix.

Budget Input tokens Output tokens Mixed tokens (70/30)
$0.01 25,000 6,250 13,158
$0.1 250,000 62,500 131,579
$0.5 1.25M 312,500 657,895
$1.00 2.5M 625,000 1.32M
$5.00 12.5M 3.13M 6.58M
$10.00 25M 6.25M 13.2M
$20.00 50M 12.5M 26.3M
$25.00 62.5M 15.6M 32.9M
$30.00 75M 18.8M 39.5M
$50.00 125M 31.3M 65.8M
$100.00 250M 62.5M 131.6M
$250.00 625M 156.3M 328.9M
$500.00 1.25B 312.5M 657.9M
$1,000.00 2.5B 625M 1.32B
$5,000.00 12.5B 3.13B 6.58B
$10,000.00 25B 6.25B 13.2B

At these rates

  • A workload with 70% input and 30% output costs $0.76 per 1 million total tokens and $76.00 per 100 million.
  • At the same token count, output costs 4 times the input rate. Response length therefore deserves its own budget limit.
  • Filling the 1,000,000-token context window with input alone would cost about $0.4, before any output.
  • Cached input is $0.04 per million, 90% below standard input when the workload meets the provider’s cache rules.

OpenAI

OpenAI API models for chat, reasoning, coding, and multimodal work.

All OpenAI models

These figures use public list prices. Cache eligibility, batch or priority modes, long-context surcharges, discounts, and regional billing can change the invoice.

Related models