200 offerings · 12 model authors

What every model costs at your volume

Leaderboards tell you which model is best. They don’t tell you what it costs to run your workload. Put your numbers in and see the spread — it’s 1200× between the cheapest and dearest model for identical work.

Cheapest
$12.50
Qwen 3.7 Flash · Alibaba
Most expensive
$2,500
GPT-5.6 sol · OpenAI
Spread
200×
same work, same month
Qwen 3.7 Flash
Alibaba · Alibaba
$12.50
GPT-5 nano
OpenAI · OpenAI
$30.00
GLM 4.7 Flash
Z.ai · Z.ai
$32.00
DeepSeek V4-Flash
DeepSeek · DeepSeek
$42.00
Mistral Small 4
Mistral · Mistral
$60.00
MiniMax M3
MiniMax · MiniMax
$120.00
Gemini 3.1 Flash-Lite
Google · Google
$125.00
Qwen 3.7 Plus
Alibaba · Alibaba
$128.00
DeepSeek V4-Pro
DeepSeek · DeepSeek
$130.50
GPT-5 mini
OpenAI · OpenAI
$150.00
Nova 2 Lite
Amazon · Amazon
$185.00
GLM 5.2
Z.ai · Z.ai
$245.50
Kimi K2.5
Moonshot AI · Moonshot AI
$256.50
Grok 4.3
xAI · xAI
$375.00
Haiku 4.5
Anthropic · Anthropic
$450.00
Qwen 3.7 Max
Alibaba · Alibaba
$516.25
Mistral Medium 3.5
Mistral · Mistral
$675.00
Gemini 3.6 Flash
Google · Google
$675.00
Grok 4.5
xAI · xAI
$700.00
Sonnet 5
Anthropic · Anthropic
$900.00
Gemini 3.1 Pro
Google · Google
$1,000
Command A
Cohere · Cohere
$1,000
GPT-5.6 terra
OpenAI · OpenAI
$1,250
Kimi K3
Moonshot AI · Moonshot AI
$1,350
Opus 5
Anthropic · Anthropic
$2,250
GPT-5.6 sol
OpenAI · OpenAI
$2,500

Showing 26 of 26 matching offerings · 12 model authors

Verified today

Where these numbers come from

Prices: published rate cards

Every figure comes from the provider’s own pricing page, stored with its source URL and the date we last confirmed it. Nothing is estimated. Where a provider tiers by prompt length we show the base tier and note the threshold.

Cost: your inputs, our arithmetic

Monthly cost is calls × (input tokens × input rate + output tokens × output rate). No hidden assumptions — change any field and the maths is visible in the result. Cache discounts are not applied, so figures are an upper bound.

Quality: we don’t publish a score

We haven’t run a cross-provider benchmark suite, so we don’t print an intelligence index. For independent quality rankings we recommend Artificial Analysis. What we do instead is benchmark your tasks against your own quality bar during an audit.

Cost alone isn’t the answer. The cheapest model that actually does your job is. Describe your task and we’ll pick it — or paste a prompt and watch it get smaller.

ProviderCached inSource
Qwen 3.7 Flash
qwen3.7-flash
Alibaba$0.03$0.131Mrate card
Nova Micro
nova-micro-v1
Amazon$0.04$0.14128Krate card
Command R7B
command-r7b-12-2024
Cohere$0.04$0.15128Krate card
GPT-5 nano
gpt-5-nano
OpenAI$0.05$0.40$0.005rate card
GLM 4.7 Flash
glm-4.7-flash
Z.ai$0.06$0.40202.752Krate card
Gemini 2.5 Flash-Lite
gemini-2.5-flash-lite
Google$0.10$0.40rate card
Qwen3 Coder Next
qwen3-coder-next
Alibaba$0.12$0.80262.144Krate card
DeepSeek V4-Flash
deepseek-v4-flash
DeepSeek$0.14$0.28$0.0031Mrate card
MiniMax M2.5
minimax-m2.5
MiniMax$0.15$0.90204.8Krate card
Mistral Small 4
mistral-small-2603
Mistral$0.15$0.60262.144Krate card
Ministral 3 8B
ministral-8b-2512
Mistral$0.15$0.15262.144Krate card
Command R
command-r-08-2024
Cohere$0.15$0.60128Krate card
Qwen 3.6 Flash
qwen3.6-flash
Alibaba$0.19$1.131Mrate card
GPT-5 mini
gpt-5-mini
OpenAI$0.25$2.00$0.025rate card
MiniMax M2.7
minimax-m2.7
MiniMax$0.25$1.00204.8Krate card
Gemini 3.1 Flash-Lite
gemini-3.1-flash-lite
Google$0.25$1.50rate card
Nova 2 Lite
nova-2-lite-v1
Amazon$0.30$2.501Mrate card
Qwen 3.6 27B
qwen3.6-27b
Alibaba$0.30$2.00262.144Krate card
MiniMax M3
minimax-m3
MiniMax$0.30$1.201.048576Mrate card
Gemini 3.5 Flash-Lite
gemini-3.5-flash-lite
Google$0.30$2.50rate card
Qwen 3.7 Plus
qwen3.7-plus
Alibaba$0.32$1.281Mrate card
Devstral 2
devstral-2512
Mistral$0.40$2.00262.144Krate card
DeepSeek V4-Pro
deepseek-v4-pro
DeepSeek$0.43$0.87$0.0041Mrate card
Mistral Large 3
mistral-large-2512
Mistral$0.50$1.50262.144Krate card
Kimi K2.5
kimi-k2.5
Moonshot AI$0.57$2.85262.144Krate card
Kimi K2.6
kimi-k2.6
Moonshot AI$0.65$2.72262.144Krate card
GLM 5.2
glm-5.2
Z.ai$0.69$2.161.048576Mrate card
Kimi K2.7 Code
kimi-k2.7-code
Moonshot AI$0.73$3.50262.144Krate card
Nova Pro
nova-pro-v1
Amazon$0.80$3.20300Krate card
GLM 5
glm-5
Z.ai$0.95$2.55204.8Krate card
GLM 5.1
glm-5.1
Z.ai$0.97$3.04204.8Krate card
Haiku 4.5
claude-haiku-4-5
Anthropic$1.00$5.00$0.100200Krate card
GPT-5.6 luna
gpt-5-6-luna
OpenAI$1.00$6.00$0.100rate card
Grok 4.3
grok-4.3
xAI$1.25$2.501Mrate card
Grok 4.20
grok-4.20
xAI$1.25$2.502Mrate card
Qwen 3.7 Max
qwen3.7-max
Alibaba$1.48$4.421Mrate card
Mistral Medium 3.5
mistral-medium-3-5
Mistral$1.50$7.50262.144Krate card
Gemini 3.6 Flash
gemini-3.6-flash
Google$1.50$7.50rate card
Sonnet 5
claude-sonnet-5
Introductory pricing through 2026-08-31; $3/$15 thereafter
Anthropic$2.00$10.00$0.2001Mrate card
Grok 4.5
grok-4.5
xAI$2.00$6.00500Krate card
Gemini 3.1 Pro
gemini-3.1-pro-preview
Above 200K input: $4.00 in / $18.00 out per MTok
Google$2.00$12.00200Krate card
Command A
command-a
Cohere$2.50$10.00256Krate card
Nova Premier
nova-premier-v1
Amazon$2.50$12.501Mrate card
Command R+
command-r-plus-08-2024
Cohere$2.50$10.00128Krate card
GPT-5.6 terra
gpt-5-6-terra
OpenAI$2.50$15.00$0.250rate card
Kimi K3
kimi-k3
Moonshot AI$3.00$15.001.048576Mrate card
GPT-5.6 sol
gpt-5-6-sol
OpenAI$5.00$30.00$0.500272Krate card
Opus 5
claude-opus-5
Anthropic$5.00$25.00$0.5001Mrate card
Fable 5
claude-fable-5
Anthropic$10.00$50.00$1.0001Mrate card
GPT-5.5 pro
gpt-5-5-pro
OpenAI$30.00$180.00272Krate card