Together AI

Gemma 3n E4B Instruct (Together) token cost

Tiny Gemma 3n E4B on Together for edge-cheap classification and routing.

Published list rates

Input

$0.06

per 1M tokens

Output

$0.12

per 1M tokens

Context window

32,000

tokens

Pricing source: Together AI pricing

Estimate cost

Estimate from text

Calculate from a token total

Project monthly volume

Daily requests × tokens per request × 30 days.

Estimated cost

Estimated tokens
Characters
Words
Input tokens
Output tokens
Input cost
Output cost
Total cost
Combined price per 1M tokens
Estimated monthly cost
Estimated daily cost

Cost as volume grows

Three views of Gemma 3n E4B Instruct (Together): sending only, receiving only, and a 70/30 mix. Same list rates, bigger rows.

Tokens Input only Output only Mix (70% in / 30% out)
1 $0.00000006 $0.00000012 $0.000000078
10 $0.0000006 $0.000001 $0.00000078
100 $0.000006 $0.000012 $0.000008
200 $0.000012 $0.000024 $0.000016
300 $0.000018 $0.000036 $0.000023
400 $0.000024 $0.000048 $0.000031
500 $0.00003 $0.00006 $0.000039
1,000 $0.00006 $0.00012 $0.000078
2,000 $0.00012 $0.00024 $0.000156
5,000 $0.0003 $0.0006 $0.00039
10,000 $0.0006 $0.0012 $0.00078
25,000 $0.0015 $0.003 $0.00195
50,000 $0.003 $0.006 $0.0039
100,000 $0.006 $0.012 $0.0078
250,000 $0.015 $0.03 $0.0195
500,000 $0.03 $0.06 $0.039
1,000,000 $0.06 $0.12 $0.078
5,000,000 $0.3 $0.6 $0.39
10,000,000 $0.6 $1.20 $0.78
50,000,000 $3.00 $6.00 $3.90
100,000,000 $6.00 $12.00 $7.80
500,000,000 $30.00 $60.00 $39.00
1,000,000,000 $60.00 $120.00 $78.00
10,000,000,000 $600.00 $1,200.00 $780.00
100,000,000,000 $6,000.00 $12,000.00 $7,800.00

Tokens for a fixed budget

Tokens Gemma 3n E4B Instruct (Together) can process for a set budget — send-only, reply-only, and a 70/30 mix.

Budget Input tokens Output tokens Mixed tokens (70/30)
$0.01 166,667 83,333 128,205
$0.1 1.67M 833,333 1.28M
$0.5 8.33M 4.17M 6.41M
$1.00 16.7M 8.33M 12.8M
$5.00 83.3M 41.7M 64.1M
$10.00 166.7M 83.3M 128.2M
$20.00 333.3M 166.7M 256.4M
$25.00 416.7M 208.3M 320.5M
$30.00 500M 250M 384.6M
$50.00 833.3M 416.7M 641M
$100.00 1.67B 833.3M 1.28B
$250.00 4.17B 2.08B 3.21B
$500.00 8.33B 4.17B 6.41B
$1,000.00 16.7B 8.33B 12.8B
$5,000.00 83.3B 41.7B 64.1B
$10,000.00 166.7B 83.3B 128.2B

At these rates

  • A workload with 70% input and 30% output costs $0.078 per 1 million total tokens and $7.80 per 100 million.
  • At the same token count, output costs 2 times the input rate. Response length therefore deserves its own budget limit.
  • Filling the 32,000-token context window with input alone would cost about $0.00192, before any output.

Together AI

Open-model serverless inference, image, audio, and dedicated GPU endpoints.

All Together AI models

These figures use public list prices. Cache eligibility, batch or priority modes, long-context surcharges, discounts, and regional billing can change the invoice.

Related models