FriendliAI models

Fast open-model Model APIs with cache-aware per-token pricing.

https://friendli.ai/pricing

FriendliAI

Fast open-model Model APIs with cache-aware per-token pricing.

There are 4 published models here. On a 70/30 mix, combined list rates run from $0.218 to $2.30 per 1 million tokens, with a median of $0.685. The largest listed context window in this group is 200,000 tokens. The most recent price check in this group is 2026-08-14.

Official FriendliAI pricing page

FriendliAI list rates

USD per million tokens. The 70/30 mix is a planning default — change it on the model page if your traffic is output-heavy.

ModelInput / 1MOutput / 1MMixed 70/30ContextVerified
Gemma 4 31B (FriendliAI) $0.14 $0.4 $0.218 128,000
MiniMax M2.5 (FriendliAI) $0.3 $1.20 $0.57 200,000
DeepSeek V3.2 (FriendliAI) $0.5 $1.50 $0.8 160,000
GLM-5.2 (FriendliAI) $1.40 $4.40 $2.30 200,000

Questions about FriendliAI rates

Do these FriendliAI figures include discounts and tax?

No. They are public list rates. Commits, credits, regions, tax, and commercial discounts are not in the number — check the source linked on each model.

When is the more expensive FriendliAI tier worth the extra?

GLM-5.2 (FriendliAI) sits near $2.30 per 1M tokens on a 70/30 mix, versus $0.218 on Gemma 4 31B (FriendliAI). Use the high tier when retries or long reasoning actually fail on the cheap one.