Cohere

Aya Expanse token cost

Multilingual Aya Expanse research models on the Cohere API at $0.50/$1.50 per 1M tokens.

Published list rates

Input

$0.5

per 1M tokens

Output

$1.50

per 1M tokens

Context window

128,000

tokens

Pricing source: Cohere pricing

Estimate cost

Estimate from text

Calculate from a token total

Project monthly volume

Daily requests × tokens per request × 30 days.

Estimated cost

Estimated tokens
Characters
Words
Input tokens
Output tokens
Input cost
Output cost
Total cost
Combined price per 1M tokens
Estimated monthly cost
Estimated daily cost

Cost as volume grows

What Aya Expanse costs if you only send tokens, only get a reply, or do both (70% in / 30% out). Rows start small and go up to huge traffic.

Tokens Input only Output only Mix (70% in / 30% out)
1 $0.0000005 $0.000002 $0.0000008
10 $0.000005 $0.000015 $0.000008
100 $0.00005 $0.00015 $0.00008
200 $0.0001 $0.0003 $0.00016
300 $0.00015 $0.00045 $0.00024
400 $0.0002 $0.0006 $0.00032
500 $0.00025 $0.00075 $0.0004
1,000 $0.0005 $0.0015 $0.0008
2,000 $0.001 $0.003 $0.0016
5,000 $0.0025 $0.0075 $0.004
10,000 $0.005 $0.015 $0.008
25,000 $0.0125 $0.0375 $0.02
50,000 $0.025 $0.075 $0.04
100,000 $0.05 $0.15 $0.08
250,000 $0.125 $0.375 $0.2
500,000 $0.25 $0.75 $0.4
1,000,000 $0.5 $1.50 $0.8
5,000,000 $2.50 $7.50 $4.00
10,000,000 $5.00 $15.00 $8.00
50,000,000 $25.00 $75.00 $40.00
100,000,000 $50.00 $150.00 $80.00
500,000,000 $250.00 $750.00 $400.00
1,000,000,000 $500.00 $1,500.00 $800.00
10,000,000,000 $5,000.00 $15,000.00 $8,000.00
100,000,000,000 $50,000.00 $150,000.00 $80,000.00

Tokens for a fixed budget

Tokens Aya Expanse can process for a set budget — send-only, reply-only, and a 70/30 mix.

Budget Input tokens Output tokens Mixed tokens (70/30)
$0.01 20,000 6,667 12,500
$0.1 200,000 66,667 125,000
$0.5 1M 333,333 625,000
$1.00 2M 666,667 1.25M
$5.00 10M 3.33M 6.25M
$10.00 20M 6.67M 12.5M
$20.00 40M 13.3M 25M
$25.00 50M 16.7M 31.3M
$30.00 60M 20M 37.5M
$50.00 100M 33.3M 62.5M
$100.00 200M 66.7M 125M
$250.00 500M 166.7M 312.5M
$500.00 1B 333.3M 625M
$1,000.00 2B 666.7M 1.25B
$5,000.00 10B 3.33B 6.25B
$10,000.00 20B 6.67B 12.5B

At these rates

  • A workload with 70% input and 30% output costs $0.8 per 1 million total tokens and $80.00 per 100 million.
  • At the same token count, output costs 3 times the input rate. Response length therefore deserves its own budget limit.
  • Filling the 128,000-token context window with input alone would cost about $0.064, before any output.

Cohere

Enterprise language models for RAG, agents, and retrieval-heavy workloads.

All Cohere models

These figures use public list prices. Cache eligibility, batch or priority modes, long-context surcharges, discounts, and regional billing can change the invoice.

Related models