Fixed billing scenario

AI API cost rankings by model

Compare the tracked token charge for one repeatable monthly workload. Use this as a starting point for budgeting, not as a claim about model quality or completed-task cost.

Fixed billing scenario

1M input + 300K output tokens per month

This board uses default source-tracked input and output rates. It excludes retries, caching discounts, tool calls, storage, and quality differences. Recalculate with your workload before routing traffic.

RankModelInput / 1MOutput / 1MMonthly referenceChecked
1GPT-OSS 20B on GroqGroq$0.07$0.30$0.162026-07-15
2Llama 4 Scout on GroqGroq$0.11$0.34$0.212026-07-15
3Gemini 2.5 Flash-LiteGoogle$0.10$0.40$0.222026-08-05
4DeepSeek V4 Flash non-thinkingDeepSeek$0.14$0.28$0.222026-08-05
5Command R 08-2024Cohere$0.15$0.60$0.332026-07-17
6GPT-4o miniOpenAI$0.15$0.60$0.332026-08-05
7GPT-OSS 120B on GroqGroq$0.15$0.60$0.332026-07-15
8Qwen3-32B on GroqGroq$0.29$0.59$0.472026-07-15
9GPT-5.6 LunaOpenAI$0.20$1.20$0.562026-08-05
10gpt-5.4-nanoOpenAI$0.20$1.25$0.572026-08-05
11DeepSeek V4 ProDeepSeek$0.43$0.87$0.702026-08-05
12Gemini 3.1 Flash-LiteGoogle$0.25$1.50$0.702026-08-05
13Gemini 3.5 Live Translate PreviewGoogle$0.25$1.50$0.702026-08-05
14Gemini 2.5 FlashGoogle$0.30$2.50$1.052026-08-05
15Gemini 3.5 Flash-LiteGoogle$0.30$2.50$1.052026-08-05
16Gemini 3 Flash PreviewGoogle$0.50$3.00$1.402026-08-05
17gpt-5.4-miniOpenAI$0.75$4.50$2.102026-08-05
18Kimi K2.6Moonshot AI$0.95$4.00$2.152026-07-17
19Kimi K2.7 CodeMoonshot AI$0.95$4.00$2.152026-07-17
20Claude Haiku 4.5Anthropic$1.00$5.00$2.502026-08-05
21Gemini 3.6 FlashGoogle$1.50$7.50$3.752026-08-05
22Mistral Medium 3.5Mistral$1.50$7.50$3.752026-07-17
23Grok 4.5xAI$2.00$6.00$3.802026-07-17
24Gemini 3.5 FlashGoogle$1.50$9.00$4.202026-08-05
25Gemini 2.5 ProGoogle$1.25$10.00$4.252026-08-05
26Kimi K2.7 Code High-SpeedMoonshot AI$1.90$8.00$4.302026-07-17
27Claude Sonnet 5Anthropic$2.00$10.00$5.002026-08-05
28Command ACohere$2.50$10.00$5.502026-07-17
29Command R+ 08-2024Cohere$2.50$10.00$5.502026-07-17
30GPT-4oOpenAI$2.50$10.00$5.502026-08-05
31Gemini 3.1 Pro PreviewGoogle$2.00$12.00$5.602026-08-05
32GPT-5.6 TerraOpenAI$2.00$12.00$5.602026-08-05
33Claude Sonnet 4.6Anthropic$3.00$15.00$7.502026-08-05
34Kimi K3Moonshot AI$3.00$15.00$7.502026-07-17
35Claude Opus 4.8Anthropic$5.00$25.00$12.502026-08-05
36Claude Opus 5Anthropic$5.00$25.00$12.502026-08-05
37GPT-5.6 SolOpenAI$5.00$30.00$14.002026-08-05
38Claude Fable 5Anthropic$10.00$50.00$25.002026-08-05
39Claude Mythos 5Anthropic$10.00$50.00$25.002026-08-05

Rates are source-tracked. Model a different workload in the LLM API cost calculator and review the provider pricing pages.

Use the result carefully

One dimension is not a complete decision

Cost, token volume, capacity, and evidence coverage answer different questions. Compare them with a representative prompt, actual output length, retry behavior, and your quality bar.

Find a model for a task
FAQ

Ranking methodology questions

Does a higher ranking mean a better model?

No. Each board measures one defined dimension. Token count, price, context capacity, and evidence coverage do not substitute for testing your own completed tasks.