Gemini 3.1 Flash-Lite token cost
Cheapest Gemini 3 Lite on the paid sheet for high-volume routing, translation, and light agent loops.

Published list rates
Input
$0.25
per 1M tokens
Output
$1.50
per 1M tokens
Cached input
$0.025
per 1M tokens
Context window
1,000,000
tokens
Pricing source: Google Gemini API pricing
Estimate cost
Estimate from text
Calculate from a token total
Project monthly volume
Daily requests × tokens per request × 30 days.
Estimated cost
- Estimated tokens
- —
- Characters
- —
- Words
- —
- Input tokens
- —
- Output tokens
- —
- Input cost
- —
- Output cost
- —
- Total cost
- —
- Combined price per 1M tokens
- —
- Estimated monthly cost
- —
- Estimated daily cost
- —
Cost as volume grows
Gemini 3.1 Flash-Lite at rising volume: prompt-only, reply-only, and a mixed chat (70% in, 30% out).
| Tokens | Input only | Output only | Mix (70% in / 30% out) |
|---|---|---|---|
| 1 | $0.00000025 | $0.000002 | $0.000000625 |
| 10 | $0.000003 | $0.000015 | $0.000006 |
| 100 | $0.000025 | $0.00015 | $0.000063 |
| 200 | $0.00005 | $0.0003 | $0.000125 |
| 300 | $0.000075 | $0.00045 | $0.000188 |
| 400 | $0.0001 | $0.0006 | $0.00025 |
| 500 | $0.000125 | $0.00075 | $0.000313 |
| 1,000 | $0.00025 | $0.0015 | $0.000625 |
| 2,000 | $0.0005 | $0.003 | $0.00125 |
| 5,000 | $0.00125 | $0.0075 | $0.003125 |
| 10,000 | $0.0025 | $0.015 | $0.00625 |
| 25,000 | $0.00625 | $0.0375 | $0.01563 |
| 50,000 | $0.0125 | $0.075 | $0.03125 |
| 100,000 | $0.025 | $0.15 | $0.0625 |
| 250,000 | $0.0625 | $0.375 | $0.15625 |
| 500,000 | $0.125 | $0.75 | $0.3125 |
| 1,000,000 | $0.25 | $1.50 | $0.625 |
| 5,000,000 | $1.25 | $7.50 | $3.125 |
| 10,000,000 | $2.50 | $15.00 | $6.25 |
| 50,000,000 | $12.50 | $75.00 | $31.25 |
| 100,000,000 | $25.00 | $150.00 | $62.50 |
| 500,000,000 | $125.00 | $750.00 | $312.50 |
| 1,000,000,000 | $250.00 | $1,500.00 | $625.00 |
| 10,000,000,000 | $2,500.00 | $15,000.00 | $6,250.00 |
| 100,000,000,000 | $25,000.00 | $150,000.00 | $62,500.00 |
Tokens for a fixed budget
If you cap spend, this is roughly how many Gemini 3.1 Flash-Lite tokens that money buys.
| Budget | Input tokens | Output tokens | Mixed tokens (70/30) |
|---|---|---|---|
| $0.01 | 40,000 | 6,667 | 16,000 |
| $0.1 | 400,000 | 66,667 | 160,000 |
| $0.5 | 2M | 333,333 | 800,000 |
| $1.00 | 4M | 666,667 | 1.6M |
| $5.00 | 20M | 3.33M | 8M |
| $10.00 | 40M | 6.67M | 16M |
| $20.00 | 80M | 13.3M | 32M |
| $25.00 | 100M | 16.7M | 40M |
| $30.00 | 120M | 20M | 48M |
| $50.00 | 200M | 33.3M | 80M |
| $100.00 | 400M | 66.7M | 160M |
| $250.00 | 1B | 166.7M | 400M |
| $500.00 | 2B | 333.3M | 800M |
| $1,000.00 | 4B | 666.7M | 1.6B |
| $5,000.00 | 20B | 3.33B | 8B |
| $10,000.00 | 40B | 6.67B | 16B |
At these rates
- A workload with 70% input and 30% output costs $0.625 per 1 million total tokens and $62.50 per 100 million.
- At the same token count, output costs 6 times the input rate. Response length therefore deserves its own budget limit.
- Filling the 1,000,000-token context window with input alone would cost about $0.25, before any output.
- Cached input is $0.025 per million, 90% below standard input when the workload meets the provider’s cache rules.
Text, image, and video input is $0.25 per 1M; audio input is $0.50 per 1M on this model.
These figures use public list prices. Cache eligibility, batch or priority modes, long-context surcharges, discounts, and regional billing can change the invoice.