One dimension is not a complete decision
Cost, token volume, capacity, and evidence coverage answer different questions. Compare them with a representative prompt, actual output length, retry behavior, and your quality bar.
Find a model for a taskCompare the tracked token charge for one repeatable monthly workload. Use this as a starting point for budgeting, not as a claim about model quality or completed-task cost.
This board uses default source-tracked input and output rates. It excludes retries, caching discounts, tool calls, storage, and quality differences. Recalculate with your workload before routing traffic.
| Rank | Model | Input / 1M | Output / 1M | Monthly reference | Checked |
|---|---|---|---|---|---|
| 1 | GPT-OSS 20B on GroqGroq | $0.07 | $0.30 | $0.16 | 2026-07-15 |
| 2 | Llama 4 Scout on GroqGroq | $0.11 | $0.34 | $0.21 | 2026-07-15 |
| 3 | Gemini 2.5 Flash-LiteGoogle | $0.10 | $0.40 | $0.22 | 2026-08-05 |
| 4 | DeepSeek V4 Flash non-thinkingDeepSeek | $0.14 | $0.28 | $0.22 | 2026-08-05 |
| 5 | Command R 08-2024Cohere | $0.15 | $0.60 | $0.33 | 2026-07-17 |
| 6 | GPT-4o miniOpenAI | $0.15 | $0.60 | $0.33 | 2026-08-05 |
| 7 | GPT-OSS 120B on GroqGroq | $0.15 | $0.60 | $0.33 | 2026-07-15 |
| 8 | Qwen3-32B on GroqGroq | $0.29 | $0.59 | $0.47 | 2026-07-15 |
| 9 | GPT-5.6 LunaOpenAI | $0.20 | $1.20 | $0.56 | 2026-08-05 |
| 10 | gpt-5.4-nanoOpenAI | $0.20 | $1.25 | $0.57 | 2026-08-05 |
| 11 | DeepSeek V4 ProDeepSeek | $0.43 | $0.87 | $0.70 | 2026-08-05 |
| 12 | Gemini 3.1 Flash-LiteGoogle | $0.25 | $1.50 | $0.70 | 2026-08-05 |
| 13 | Gemini 3.5 Live Translate PreviewGoogle | $0.25 | $1.50 | $0.70 | 2026-08-05 |
| 14 | Gemini 2.5 FlashGoogle | $0.30 | $2.50 | $1.05 | 2026-08-05 |
| 15 | Gemini 3.5 Flash-LiteGoogle | $0.30 | $2.50 | $1.05 | 2026-08-05 |
| 16 | Gemini 3 Flash PreviewGoogle | $0.50 | $3.00 | $1.40 | 2026-08-05 |
| 17 | gpt-5.4-miniOpenAI | $0.75 | $4.50 | $2.10 | 2026-08-05 |
| 18 | Kimi K2.6Moonshot AI | $0.95 | $4.00 | $2.15 | 2026-07-17 |
| 19 | Kimi K2.7 CodeMoonshot AI | $0.95 | $4.00 | $2.15 | 2026-07-17 |
| 20 | Claude Haiku 4.5Anthropic | $1.00 | $5.00 | $2.50 | 2026-08-05 |
| 21 | Gemini 3.6 FlashGoogle | $1.50 | $7.50 | $3.75 | 2026-08-05 |
| 22 | Mistral Medium 3.5Mistral | $1.50 | $7.50 | $3.75 | 2026-07-17 |
| 23 | Grok 4.5xAI | $2.00 | $6.00 | $3.80 | 2026-07-17 |
| 24 | Gemini 3.5 FlashGoogle | $1.50 | $9.00 | $4.20 | 2026-08-05 |
| 25 | Gemini 2.5 ProGoogle | $1.25 | $10.00 | $4.25 | 2026-08-05 |
| 26 | Kimi K2.7 Code High-SpeedMoonshot AI | $1.90 | $8.00 | $4.30 | 2026-07-17 |
| 27 | Claude Sonnet 5Anthropic | $2.00 | $10.00 | $5.00 | 2026-08-05 |
| 28 | Command ACohere | $2.50 | $10.00 | $5.50 | 2026-07-17 |
| 29 | Command R+ 08-2024Cohere | $2.50 | $10.00 | $5.50 | 2026-07-17 |
| 30 | GPT-4oOpenAI | $2.50 | $10.00 | $5.50 | 2026-08-05 |
| 31 | Gemini 3.1 Pro PreviewGoogle | $2.00 | $12.00 | $5.60 | 2026-08-05 |
| 32 | GPT-5.6 TerraOpenAI | $2.00 | $12.00 | $5.60 | 2026-08-05 |
| 33 | Claude Sonnet 4.6Anthropic | $3.00 | $15.00 | $7.50 | 2026-08-05 |
| 34 | Kimi K3Moonshot AI | $3.00 | $15.00 | $7.50 | 2026-07-17 |
| 35 | Claude Opus 4.8Anthropic | $5.00 | $25.00 | $12.50 | 2026-08-05 |
| 36 | Claude Opus 5Anthropic | $5.00 | $25.00 | $12.50 | 2026-08-05 |
| 37 | GPT-5.6 SolOpenAI | $5.00 | $30.00 | $14.00 | 2026-08-05 |
| 38 | Claude Fable 5Anthropic | $10.00 | $50.00 | $25.00 | 2026-08-05 |
| 39 | Claude Mythos 5Anthropic | $10.00 | $50.00 | $25.00 | 2026-08-05 |
Rates are source-tracked. Model a different workload in the LLM API cost calculator and review the provider pricing pages.
Cost, token volume, capacity, and evidence coverage answer different questions. Compare them with a representative prompt, actual output length, retry behavior, and your quality bar.
Find a model for a taskNo. Each board measures one defined dimension. Token count, price, context capacity, and evidence coverage do not substitute for testing your own completed tasks.