Microsoft

MAI-Code-1.1-Flash token cost

Cheap Microsoft coding workhorse used in GitHub Copilot, at $0.20/$1.20 per 1M with 256K context and native vision.

Published list rates

Input

$0.2

per 1M tokens

Output

$1.20

per 1M tokens

Context window

256,000

tokens

Pricing source: Microsoft MAI-Code-1.1-Flash pricing

Estimate cost

Estimate from text

Calculate from a token total

Project monthly volume

Daily requests × tokens per request × 30 days.

Estimated cost

Estimated tokens
Characters
Words
Input tokens
Output tokens
Input cost
Output cost
Total cost
Combined price per 1M tokens
Estimated monthly cost
Estimated daily cost

Cost as volume grows

Read down to see MAI-Code-1.1-Flash get more expensive as tokens grow. The last column is a typical mix: 70% input, 30% output.

Tokens Input only Output only Mix (70% in / 30% out)
1 $0.0000002 $0.000001 $0.0000005
10 $0.000002 $0.000012 $0.000005
100 $0.00002 $0.00012 $0.00005
200 $0.00004 $0.00024 $0.0001
300 $0.00006 $0.00036 $0.00015
400 $0.00008 $0.00048 $0.0002
500 $0.0001 $0.0006 $0.00025
1,000 $0.0002 $0.0012 $0.0005
2,000 $0.0004 $0.0024 $0.001
5,000 $0.001 $0.006 $0.0025
10,000 $0.002 $0.012 $0.005
25,000 $0.005 $0.03 $0.0125
50,000 $0.01 $0.06 $0.025
100,000 $0.02 $0.12 $0.05
250,000 $0.05 $0.3 $0.125
500,000 $0.1 $0.6 $0.25
1,000,000 $0.2 $1.20 $0.5
5,000,000 $1.00 $6.00 $2.50
10,000,000 $2.00 $12.00 $5.00
50,000,000 $10.00 $60.00 $25.00
100,000,000 $20.00 $120.00 $50.00
500,000,000 $100.00 $600.00 $250.00
1,000,000,000 $200.00 $1,200.00 $500.00
10,000,000,000 $2,000.00 $12,000.00 $5,000.00
100,000,000,000 $20,000.00 $120,000.00 $50,000.00

Tokens for a fixed budget

Tokens MAI-Code-1.1-Flash can process for a set budget — send-only, reply-only, and a 70/30 mix.

Budget Input tokens Output tokens Mixed tokens (70/30)
$0.01 50,000 8,333 20,000
$0.1 500,000 83,333 200,000
$0.5 2.5M 416,667 1M
$1.00 5M 833,333 2M
$5.00 25M 4.17M 10M
$10.00 50M 8.33M 20M
$20.00 100M 16.7M 40M
$25.00 125M 20.8M 50M
$30.00 150M 25M 60M
$50.00 250M 41.7M 100M
$100.00 500M 83.3M 200M
$250.00 1.25B 208.3M 500M
$500.00 2.5B 416.7M 1B
$1,000.00 5B 833.3M 2B
$5,000.00 25B 4.17B 10B
$10,000.00 50B 8.33B 20B

At these rates

  • A workload with 70% input and 30% output costs $0.5 per 1 million total tokens and $50.00 per 100 million.
  • At the same token count, output costs 6 times the input rate. Response length therefore deserves its own budget limit.
  • Filling the 256,000-token context window with input alone would cost about $0.0512, before any output.

Microsoft

First-party Microsoft AI (MAI) reasoning and coding models on Microsoft Foundry.

All Microsoft models

These figures use public list prices. Cache eligibility, batch or priority modes, long-context surcharges, discounts, and regional billing can change the invoice.

Related models