Groq

Llama Prompt Guard 2 86M (Groq) token cost

Larger Llama Prompt Guard 2 86M on Groq at $0.04/$0.04 when you want more classifier headroom than 22M.

Published list rates

Input

$0.04

per 1M tokens

Output

$0.04

per 1M tokens

Context window

512

tokens

Pricing source: GroqCloud models

Estimate cost

Estimate from text

Calculate from a token total

Project monthly volume

Daily requests × tokens per request × 30 days.

Estimated cost

Estimated tokens
Characters
Words
Input tokens
Output tokens
Input cost
Output cost
Total cost
Combined price per 1M tokens
Estimated monthly cost
Estimated daily cost

Cost as volume grows

What Llama Prompt Guard 2 86M (Groq) costs if you only send tokens, only get a reply, or do both (70% in / 30% out). Rows start small and go up to huge traffic.

Tokens Input only Output only Mix (70% in / 30% out)
1 $0.00000004 $0.00000004 $0.00000004
10 $0.0000004 $0.0000004 $0.0000004
100 $0.000004 $0.000004 $0.000004
200 $0.000008 $0.000008 $0.000008
300 $0.000012 $0.000012 $0.000012
400 $0.000016 $0.000016 $0.000016
500 $0.00002 $0.00002 $0.00002
1,000 $0.00004 $0.00004 $0.00004
2,000 $0.00008 $0.00008 $0.00008
5,000 $0.0002 $0.0002 $0.0002
10,000 $0.0004 $0.0004 $0.0004
25,000 $0.001 $0.001 $0.001
50,000 $0.002 $0.002 $0.002
100,000 $0.004 $0.004 $0.004
250,000 $0.01 $0.01 $0.01
500,000 $0.02 $0.02 $0.02
1,000,000 $0.04 $0.04 $0.04
5,000,000 $0.2 $0.2 $0.2
10,000,000 $0.4 $0.4 $0.4
50,000,000 $2.00 $2.00 $2.00
100,000,000 $4.00 $4.00 $4.00
500,000,000 $20.00 $20.00 $20.00
1,000,000,000 $40.00 $40.00 $40.00
10,000,000,000 $400.00 $400.00 $400.00
100,000,000,000 $4,000.00 $4,000.00 $4,000.00

Tokens for a fixed budget

Tokens Llama Prompt Guard 2 86M (Groq) can process for a set budget — send-only, reply-only, and a 70/30 mix.

Budget Input tokens Output tokens Mixed tokens (70/30)
$0.01 250,000 250,000 250,000
$0.1 2.5M 2.5M 2.5M
$0.5 12.5M 12.5M 12.5M
$1.00 25M 25M 25M
$5.00 125M 125M 125M
$10.00 250M 250M 250M
$20.00 500M 500M 500M
$25.00 625M 625M 625M
$30.00 750M 750M 750M
$50.00 1.25B 1.25B 1.25B
$100.00 2.5B 2.5B 2.5B
$250.00 6.25B 6.25B 6.25B
$500.00 12.5B 12.5B 12.5B
$1,000.00 25B 25B 25B
$5,000.00 125B 125B 125B
$10,000.00 250B 250B 250B

At these rates

  • A workload with 70% input and 30% output costs $0.04 per 1 million total tokens and $4.00 per 100 million.
  • At the same token count, output costs 1 times the input rate. Response length therefore deserves its own budget limit.
  • Filling the 512-token context window with input alone would cost about $0.00002, before any output.

Groq

Ultra-fast inference for open models with transparent per-token pricing.

All Groq models

These figures use public list prices. Cache eligibility, batch or priority modes, long-context surcharges, discounts, and regional billing can change the invoice.

Related models