Cloudflare
Llama 3.1 8B Instruct FP8 Fast (Workers AI) token cost
Meta Llama 3.1 8B FP8 Fast on Cloudflare Workers AI at $0.045 / $0.384 per 1M tokens (neuron-backed list rates).

Published list rates
Input
$0.045
per 1M tokens
Output
$0.384
per 1M tokens
Context window
128,000
tokens
Pricing source: Cloudflare Workers AI pricing
Estimate cost
Estimate from text
Calculate from a token total
Project monthly volume
Daily requests × tokens per request × 30 days.
Estimated cost
- Estimated tokens
- —
- Characters
- —
- Words
- —
- Input tokens
- —
- Output tokens
- —
- Input cost
- —
- Output cost
- —
- Total cost
- —
- Combined price per 1M tokens
- —
- Estimated monthly cost
- —
- Estimated daily cost
- —
Cost as volume grows
Read down to see Llama 3.1 8B Instruct FP8 Fast (Workers AI) get more expensive as tokens grow. The last column is a typical mix: 70% input, 30% output.
| Tokens | Input only | Output only | Mix (70% in / 30% out) |
|---|---|---|---|
| 1 | $0.000000045 | $0.000000384 | $0.0000001467 |
| 10 | $0.00000045 | $0.000004 | $0.000001 |
| 100 | $0.000005 | $0.000038 | $0.000015 |
| 200 | $0.000009 | $0.000077 | $0.000029 |
| 300 | $0.000013 | $0.000115 | $0.000044 |
| 400 | $0.000018 | $0.000154 | $0.000059 |
| 500 | $0.000022 | $0.000192 | $0.000073 |
| 1,000 | $0.000045 | $0.000384 | $0.000147 |
| 2,000 | $0.00009 | $0.000768 | $0.000293 |
| 5,000 | $0.000225 | $0.00192 | $0.000734 |
| 10,000 | $0.00045 | $0.00384 | $0.001467 |
| 25,000 | $0.001125 | $0.0096 | $0.003667 |
| 50,000 | $0.00225 | $0.0192 | $0.007335 |
| 100,000 | $0.0045 | $0.0384 | $0.01467 |
| 250,000 | $0.01125 | $0.096 | $0.03668 |
| 500,000 | $0.0225 | $0.192 | $0.07335 |
| 1,000,000 | $0.045 | $0.384 | $0.1467 |
| 5,000,000 | $0.225 | $1.92 | $0.7335 |
| 10,000,000 | $0.45 | $3.84 | $1.467 |
| 50,000,000 | $2.25 | $19.20 | $7.335 |
| 100,000,000 | $4.50 | $38.40 | $14.67 |
| 500,000,000 | $22.50 | $192.00 | $73.35 |
| 1,000,000,000 | $45.00 | $384.00 | $146.70 |
| 10,000,000,000 | $450.00 | $3,840.00 | $1,467.00 |
| 100,000,000,000 | $4,500.00 | $38,400.00 | $14,670.00 |
Tokens for a fixed budget
Tokens Llama 3.1 8B Instruct FP8 Fast (Workers AI) can process for a set budget — send-only, reply-only, and a 70/30 mix.
| Budget | Input tokens | Output tokens | Mixed tokens (70/30) |
|---|---|---|---|
| $0.01 | 222,222 | 26,042 | 68,166 |
| $0.1 | 2.22M | 260,417 | 681,663 |
| $0.5 | 11.1M | 1.3M | 3.41M |
| $1.00 | 22.2M | 2.6M | 6.82M |
| $5.00 | 111.1M | 13M | 34.1M |
| $10.00 | 222.2M | 26M | 68.2M |
| $20.00 | 444.4M | 52.1M | 136.3M |
| $25.00 | 555.6M | 65.1M | 170.4M |
| $30.00 | 666.7M | 78.1M | 204.5M |
| $50.00 | 1.11B | 130.2M | 340.8M |
| $100.00 | 2.22B | 260.4M | 681.7M |
| $250.00 | 5.56B | 651M | 1.7B |
| $500.00 | 11.1B | 1.3B | 3.41B |
| $1,000.00 | 22.2B | 2.6B | 6.82B |
| $5,000.00 | 111.1B | 13B | 34.1B |
| $10,000.00 | 222.2B | 26B | 68.2B |
At these rates
- A workload with 70% input and 30% output costs $0.1467 per 1 million total tokens and $14.67 per 100 million.
- At the same token count, output costs 8.53 times the input rate. Response length therefore deserves its own budget limit.
- Filling the 128,000-token context window with input alone would cost about $0.00576, before any output.
Workers AI bills via Neurons ($0.011 per 1,000 Neurons) with a free daily allocation. Token rates above are Cloudflare’s published model-equivalent prices.
These figures use public list prices. Cache eligibility, batch or priority modes, long-context surcharges, discounts, and regional billing can change the invoice.