Fixed billing scenario

AI API cost rankings by model

Compare the tracked token charge for one repeatable monthly workload. Use this as a starting point for budgeting, not as a claim about model quality or completed-task cost.

Fixed billing scenario

1M input + 300K output tokens per month

This board uses default source-tracked input and output rates. It excludes retries, caching discounts, tool calls, storage, and quality differences. Recalculate with your workload before routing traffic.

RankModelInput / 1MOutput / 1MMonthly referenceChecked
1GPT-OSS 20B on GroqGroq$0.07$0.30$0.162026-07-15
2Llama 4 Scout on GroqGroq$0.11$0.34$0.212026-07-15
3Gemini 2.5 Flash-LiteGoogle$0.10$0.40$0.222026-08-26
4DeepSeek V4 Flash non-thinkingDeepSeek$0.14$0.28$0.222026-08-16
5Command R 08-2024Cohere$0.15$0.60$0.332026-07-17
6GPT-4o miniOpenAI$0.15$0.60$0.332026-09-08
7GPT-OSS 120B on GroqGroq$0.15$0.60$0.332026-07-15
8Qwen3-32B on GroqGroq$0.29$0.59$0.472026-07-15
9GPT-5.6 LunaOpenAI$0.20$1.20$0.562026-09-08
10gpt-5.4-nanoOpenAI$0.20$1.25$0.572026-09-08
11DeepSeek V4 ProDeepSeek$0.43$0.87$0.702026-08-16
12Gemini 3.1 Flash-LiteGoogle$0.25$1.50$0.702026-08-26
13Gemini 3.5 Live Translate PreviewGoogle$0.25$1.50$0.702026-08-26
14Gemini 2.5 FlashGoogle$0.30$2.50$1.052026-08-26
15Gemini 3.5 Flash-LiteGoogle$0.30$2.50$1.052026-08-26
16Gemini 3 Flash PreviewGoogle$0.50$3.00$1.402026-08-26
17Gemini 3.6 FlashGoogle$0.75$3.75$1.882026-09-03
18Gemini 3.7 FlashGoogle$0.75$3.75$1.882026-09-03
19Gemini 3.8 FlashGoogle$0.75$3.75$1.882026-09-03
20gpt-5.4-miniOpenAI$0.75$4.50$2.102026-09-08
21Kimi K2.6Moonshot AI$0.95$4.00$2.152026-07-17
22Kimi K2.7 CodeMoonshot AI$0.95$4.00$2.152026-07-17
23Claude Haiku 4.5Anthropic$1.00$5.00$2.502026-09-08
24Mistral Medium 3.5Mistral$1.50$7.50$3.752026-07-17
25Grok 4.5xAI$2.00$6.00$3.802026-07-17
26Gemini 3.5 FlashGoogle$1.50$9.00$4.202026-08-26
27Gemini 2.5 ProGoogle$1.25$10.00$4.252026-08-26
28Kimi K2.7 Code High-SpeedMoonshot AI$1.90$8.00$4.302026-07-17
29Claude Sonnet 5Anthropic$2.00$10.00$5.002026-09-08
30Command ACohere$2.50$10.00$5.502026-07-17
31Command R+ 08-2024Cohere$2.50$10.00$5.502026-07-17
32GPT-4oOpenAI$2.50$10.00$5.502026-09-08
33Gemini 3.1 Pro PreviewGoogle$2.00$12.00$5.602026-08-26
34GPT-5.6 TerraOpenAI$2.00$12.00$5.602026-09-08
35Claude Sonnet 4.6Anthropic$3.00$15.00$7.502026-09-08
36Kimi K3Moonshot AI$3.00$15.00$7.502026-07-17
37GPT-5.6 SolOpenAI$4.00$20.00$10.002026-09-08
38Claude Opus 4.8Anthropic$5.00$25.00$12.502026-09-08
39Claude Opus 5Anthropic$5.00$25.00$12.502026-09-08
40Claude Fable 5Anthropic$10.00$50.00$25.002026-09-08

Rates are source-tracked. Model a different workload in the LLM API cost calculator and review the provider pricing pages.

Use the result carefully

One dimension is not a complete decision

Cost, token volume, capacity, and evidence coverage answer different questions. Compare them with a representative prompt, actual output length, retry behavior, and your quality bar.

Find a model for a task
FAQ

Ranking methodology questions

Does a higher ranking mean a better model?

No. Each board measures one defined dimension. Token count, price, context capacity, and evidence coverage do not substitute for testing your own completed tasks.