Parasail

Gemma 4 31B (Parasail) token cost

Gemma 4 31B on Parasail for fast open multimodal-class inference.

Published list rates

Input

$0.15

per 1M tokens

Output

$0.4

per 1M tokens

Cached input

$0.06

per 1M tokens

Context window

128,000

tokens

Pricing source: Parasail pricing

Estimate cost

Estimate from text

Calculate from a token total

Project monthly volume

Daily requests × tokens per request × 30 days.

Estimated cost

Estimated tokens
Characters
Words
Input tokens
Output tokens
Input cost
Output cost
Total cost
Combined price per 1M tokens
Estimated monthly cost
Estimated daily cost

Cost as volume grows

What Gemma 4 31B (Parasail) costs if you only send tokens, only get a reply, or do both (70% in / 30% out). Rows start small and go up to huge traffic.

Tokens Input only Output only Mix (70% in / 30% out)
1 $0.00000015 $0.0000004 $0.000000225
10 $0.000002 $0.000004 $0.000002
100 $0.000015 $0.00004 $0.000022
200 $0.00003 $0.00008 $0.000045
300 $0.000045 $0.00012 $0.000068
400 $0.00006 $0.00016 $0.00009
500 $0.000075 $0.0002 $0.000113
1,000 $0.00015 $0.0004 $0.000225
2,000 $0.0003 $0.0008 $0.00045
5,000 $0.00075 $0.002 $0.001125
10,000 $0.0015 $0.004 $0.00225
25,000 $0.00375 $0.01 $0.005625
50,000 $0.0075 $0.02 $0.01125
100,000 $0.015 $0.04 $0.0225
250,000 $0.0375 $0.1 $0.05625
500,000 $0.075 $0.2 $0.1125
1,000,000 $0.15 $0.4 $0.225
5,000,000 $0.75 $2.00 $1.125
10,000,000 $1.50 $4.00 $2.25
50,000,000 $7.50 $20.00 $11.25
100,000,000 $15.00 $40.00 $22.50
500,000,000 $75.00 $200.00 $112.50
1,000,000,000 $150.00 $400.00 $225.00
10,000,000,000 $1,500.00 $4,000.00 $2,250.00
100,000,000,000 $15,000.00 $40,000.00 $22,500.00

Tokens for a fixed budget

How far a fixed spend goes on Gemma 4 31B (Parasail) at these list rates.

Budget Input tokens Output tokens Mixed tokens (70/30)
$0.01 66,667 25,000 44,444
$0.1 666,667 250,000 444,444
$0.5 3.33M 1.25M 2.22M
$1.00 6.67M 2.5M 4.44M
$5.00 33.3M 12.5M 22.2M
$10.00 66.7M 25M 44.4M
$20.00 133.3M 50M 88.9M
$25.00 166.7M 62.5M 111.1M
$30.00 200M 75M 133.3M
$50.00 333.3M 125M 222.2M
$100.00 666.7M 250M 444.4M
$250.00 1.67B 625M 1.11B
$500.00 3.33B 1.25B 2.22B
$1,000.00 6.67B 2.5B 4.44B
$5,000.00 33.3B 12.5B 22.2B
$10,000.00 66.7B 25B 44.4B

At these rates

  • A workload with 70% input and 30% output costs $0.225 per 1 million total tokens and $22.50 per 100 million.
  • At the same token count, output costs 2.67 times the input rate. Response length therefore deserves its own budget limit.
  • Filling the 128,000-token context window with input alone would cost about $0.0192, before any output.
  • Cached input is $0.06 per million, 60% below standard input when the workload meets the provider’s cache rules.

Parasail

Serverless open-model inference with public per-token rates and cache discounts.

All Parasail models

These figures use public list prices. Cache eligibility, batch or priority modes, long-context surcharges, discounts, and regional billing can change the invoice.

Related models