Novita AI

Llama 3.1 8B Instruct (Novita) token cost

Llama 3.1 8B on Novita at $0.02/$0.05 — budget host for classification and light agents.

Published list rates

Input

$0.02

per 1M tokens

Output

$0.05

per 1M tokens

Context window

16,384

tokens

Pricing source: Novita AI pricing

Estimate cost

Estimate from text

Calculate from a token total

Project monthly volume

Daily requests × tokens per request × 30 days.

Estimated cost

Estimated tokens
Characters
Words
Input tokens
Output tokens
Input cost
Output cost
Total cost
Combined price per 1M tokens
Estimated monthly cost
Estimated daily cost

Cost as volume grows

What Llama 3.1 8B Instruct (Novita) costs if you only send tokens, only get a reply, or do both (70% in / 30% out). Rows start small and go up to huge traffic.

Tokens Input only Output only Mix (70% in / 30% out)
1 $0.00000002 $0.00000005 $0.000000029
10 $0.0000002 $0.0000005 $0.00000029
100 $0.000002 $0.000005 $0.000003
200 $0.000004 $0.00001 $0.000006
300 $0.000006 $0.000015 $0.000009
400 $0.000008 $0.00002 $0.000012
500 $0.00001 $0.000025 $0.000015
1,000 $0.00002 $0.00005 $0.000029
2,000 $0.00004 $0.0001 $0.000058
5,000 $0.0001 $0.00025 $0.000145
10,000 $0.0002 $0.0005 $0.00029
25,000 $0.0005 $0.00125 $0.000725
50,000 $0.001 $0.0025 $0.00145
100,000 $0.002 $0.005 $0.0029
250,000 $0.005 $0.0125 $0.00725
500,000 $0.01 $0.025 $0.0145
1,000,000 $0.02 $0.05 $0.029
5,000,000 $0.1 $0.25 $0.145
10,000,000 $0.2 $0.5 $0.29
50,000,000 $1.00 $2.50 $1.45
100,000,000 $2.00 $5.00 $2.90
500,000,000 $10.00 $25.00 $14.50
1,000,000,000 $20.00 $50.00 $29.00
10,000,000,000 $200.00 $500.00 $290.00
100,000,000,000 $2,000.00 $5,000.00 $2,900.00

Tokens for a fixed budget

How far a fixed spend goes on Llama 3.1 8B Instruct (Novita) at these list rates.

Budget Input tokens Output tokens Mixed tokens (70/30)
$0.01 500,000 200,000 344,828
$0.1 5M 2M 3.45M
$0.5 25M 10M 17.2M
$1.00 50M 20M 34.5M
$5.00 250M 100M 172.4M
$10.00 500M 200M 344.8M
$20.00 1B 400M 689.7M
$25.00 1.25B 500M 862.1M
$30.00 1.5B 600M 1.03B
$50.00 2.5B 1B 1.72B
$100.00 5B 2B 3.45B
$250.00 12.5B 5B 8.62B
$500.00 25B 10B 17.2B
$1,000.00 50B 20B 34.5B
$5,000.00 250B 100B 172.4B
$10,000.00 500B 200B 344.8B

At these rates

  • A workload with 70% input and 30% output costs $0.029 per 1 million total tokens and $2.90 per 100 million.
  • At the same token count, output costs 2.5 times the input rate. Response length therefore deserves its own budget limit.
  • Filling the 16,384-token context window with input alone would cost about $0.000328, before any output.

Novita AI

Open-model serverless API with competitive per-token rates across Llama, Qwen, and DeepSeek.

All Novita AI models

These figures use public list prices. Cache eligibility, batch or priority modes, long-context surcharges, discounts, and regional billing can change the invoice.

Related models