Cloudflare

Llama 3.1 8B Instruct FP8 Fast (Workers AI) token cost calculator

Meta Llama 3.1 8B FP8 Fast on Cloudflare Workers AI at $0.045 / $0.384 per 1M tokens (neuron-backed list rates).

Published list rates

Input

$0.045

per 1M tokens

Output

$0.384

per 1M tokens

Context window

128,000

tokens

Pricing source: Cloudflare Workers AI pricing

Token cost calculator

List prices are stored in US dollars. Conversions use the units-per-dollar exchange-rate snapshot last checked 2026-07-31

Estimate from text

Calculate from a token total

Project monthly volume

Uses the input/output mix above and a 30-day month: requests per day × tokens per request × 30.

Estimated cost

Estimated tokens
Characters
Words
Input tokens
Output tokens
Input cost
Output cost
Total cost
Combined price per 1M tokens
Estimated monthly cost
Estimated daily cost

Cost by token volume

Compare input-only, output-only, and 70/30 mixed workloads from one token to 100 billion. Intermediate bands make smaller production steps easier to price.

Tokens Input only Output only 70/30 mix
1 $0.000000045 $0.000000384 $0.0000001467
10 $0.00000045 $0.000004 $0.000001
100 $0.000005 $0.000038 $0.000015
200 $0.000009 $0.000077 $0.000029
300 $0.000013 $0.000115 $0.000044
400 $0.000018 $0.000154 $0.000059
500 $0.000022 $0.000192 $0.000073
1,000 $0.000045 $0.000384 $0.000147
2,000 $0.00009 $0.000768 $0.000293
5,000 $0.000225 $0.00192 $0.000734
10,000 $0.00045 $0.00384 $0.001467
25,000 $0.001125 $0.0096 $0.003667
50,000 $0.00225 $0.0192 $0.007335
100,000 $0.0045 $0.0384 $0.01467
250,000 $0.01125 $0.096 $0.03668
500,000 $0.0225 $0.192 $0.07335
1,000,000 $0.045 $0.384 $0.1467
5,000,000 $0.225 $1.92 $0.7335
10,000,000 $0.45 $3.84 $1.467
50,000,000 $2.25 $19.20 $7.335
100,000,000 $4.50 $38.40 $14.67
500,000,000 $22.50 $192.00 $73.35
1,000,000,000 $45.00 $384.00 $146.70
10,000,000,000 $450.00 $3,840.00 $1,467.00
100,000,000,000 $4,500.00 $38,400.00 $14,670.00

Tokens available by budget

See how many tokens a fixed budget buys at the published rates. Display-currency values use the current exchange-rate snapshot shown on the page.

Budget Input tokens Output tokens Mixed tokens (70/30)
$0.01 222,222 26,042 68,166
$0.1 2.22M 260,417 681,663
$0.5 11.1M 1.3M 3.41M
$1.00 22.2M 2.6M 6.82M
$5.00 111.1M 13M 34.1M
$10.00 222.2M 26M 68.2M
$20.00 444.4M 52.1M 136.3M
$25.00 555.6M 65.1M 170.4M
$30.00 666.7M 78.1M 204.5M
$50.00 1.11B 130.2M 340.8M
$100.00 2.22B 260.4M 681.7M
$250.00 5.56B 651M 1.7B
$500.00 11.1B 1.3B 3.41B
$1,000.00 22.2B 2.6B 6.82B
$5,000.00 111.1B 13B 34.1B
$10,000.00 222.2B 26B 68.2B

Cost profile

  • A workload with 70% input and 30% output costs $0.1467 per 1 million total tokens and $14.67 per 100 million.
  • At the same token count, output costs 8.53 times the input rate. Response length therefore deserves its own budget limit.
  • Filling the 128,000-token context window with input alone would cost about $0.00576, before any output.

About Cloudflare

Edge compute and storage products with usage-based add-ons. The Llama 3.1 8B Instruct FP8 Fast (Workers AI) page covers this meter in detail; the Cloudflare hub gathers the rest of the lineup with a pricing table, list-rate ranges, and calculators side by side.

See all Cloudflare models and the pricing table

How the text estimate works

For rough planning, CostUse uses about four characters per token and about 0.75 words per token. Actual tokenization changes with language, model, formatting, and special tokens.

Workers AI bills via Neurons ($0.011 per 1,000 Neurons) with a free daily allocation. Token rates above are Cloudflare’s published model-equivalent prices.

These results use public list prices. Cache eligibility, batch or priority modes, long-context surcharges, discounts, and regional billing can change the final invoice.

Questions about this rate

How much is 1 million mixed tokens on Llama 3.1 8B Instruct FP8 Fast (Workers AI)?

About $0.1467 at a 70% input / 30% output mix using list prices verified 2026-07-30.

How much are 100 million mixed tokens?

About $14.67 at the same 70/30 mix — useful for rough monthly volume planning.

What is the Llama 3.1 8B Instruct FP8 Fast (Workers AI) price per million input and output tokens?

Input $0.045 and output $0.384 per 1M tokens (USD list rates). Use the calculator above with your real token mix.

Is the text-to-token estimate exact?

No. The estimator approximates tokens from characters and words. Real bills depend on the provider tokenizer, special tokens, and discounts (cache, batch, priority).

What does cost per token mean, and why is output usually more expensive?

Providers charge for tokens processed. Output tokens (generated replies) often cost several times more than input tokens (your prompt), so completion length can dominate the bill.

Related models