The same answer costs 2,000 times more on one model than another.
Providers publish price per million tokens, which is not comparable across models. This index prices one standard answer, 1,000 input tokens plus 500 output tokens, across 126 models from 33 providers. Last updated August 20, 2026.
Headline figures
Cheapest answer: $0.000053 on Ling-3.0-flash (inclusionAI).
Median answer: $0.0019 on Kimi K2 Thinking.
Most expensive answer: $0.105 on Claude Opus 4.7 (Fast) (Anthropic).
Spread from cheapest to most expensive: about 2,000 times.
How this is calculated
One standard answer is 1,000 input tokens, roughly a prompt with a couple of paragraphs of context, and 500 output tokens, roughly a substantial reply. Every model is priced on that identical workload: (input rate x 1,000 + output rate x 500) divided by 1,000,000.
Change the mix and the ordering shifts, which is the point: a model with cheap input and dear output wins on short prompts and loses on long replies. The formula is published so you can recompute it against your own mix.
Sources and limits
Prices are public list rates from the OpenRouter model catalogue, in USD per million tokens, fetched August 20, 2026.
They exclude prompt caching discounts, batch pricing, and negotiated volume rates. Real bills at scale run lower.
Cost is not quality. A model 2,000 times cheaper is not 2,000 times worse, and on many everyday tasks the gap is small.
The credits column is what Whizi charges per message on its own plans, shown for the 115 models it meters individually. It is not a provider rate.
The full index
126 models, cheapest first. Free to reuse with attribution to the Whizi AI Model Cost Index and a link to this page. Raw data: JSON and CSV.
#
Model
Provider
Per answer
Per 1,000 answers
$ / 1M in
$ / 1M out
Context
Whizi credits
1
Ling-3.0-flash
inclusionAI
$0.000053
$0.052
$0.021
$0.063
262K
1x
2
Nex-N2-Mini
Nex Agi
$0.000075
$0.075
$0.025
$0.1
262K
1x
3
Solar Pro 4
Upstage
$0.000090
$0.090
$0.03
$0.12
524K
1x
4
Qwen3.7 Flash
Qwen
$0.000095
$0.095
$0.03
$0.13
1M
1x
5
Granite 4.1 8B
Ibm Granite
$0.000100
$0.100
$0.05
$0.1
131K
1x
6
Mercury 2.5
Inception
$0.000115
$0.115
$0.04
$0.15
260K
1x
7
Laguna XS 2.1
Poolside
$0.000120
$0.120
$0.06
$0.12
262K
1x
8
Hy-MT2-1.8B
Tencent
$0.000133
$0.133
$0.044
$0.177
8K
not metered
9
Phi 4
Microsoft
$0.000140
$0.140
$0.07
$0.14
16K
1x
10
Solar Mini 4
Upstage
$0.000150
$0.150
$0.05
$0.2
524K
1x
11
Nemotron 3.5 Lightning
NVIDIA
$0.000180
$0.180
$0.08
$0.2
262K
1x
12
Laguna S 2.1
Poolside
$0.000180
$0.180
$0.09
$0.18
1M
1x
13
Granite 4.2 8B
IBM
$0.000185
$0.185
$0.06
$0.25
131K
1x
14
Hy-MT2-30B-A3B
Tencent
$0.000222
$0.222
$0.074
$0.295
8K
not metered
15
Llama 3.3 70B Instruct
Meta
$0.000260
$0.260
$0.1
$0.32
131K
1x
16
DeepSeek V4 Flash 0731
DeepSeek
$0.000280
$0.280
$0.14
$0.28
1M
1x
17
MiMo-V2.6-Flash
Xiaomi
$0.000280
$0.280
$0.14
$0.28
1M
1x
18
GLM 5.3 Flash
Z.ai
$0.000290
$0.290
$0.04
$0.5
1M
1x
19
Gemini 2.5 Flash Lite
Google
$0.000300
$0.300
$0.1
$0.4
1M
1x
20
GPT-6 Luna
OpenAI
$0.000350
$0.350
$0.1
$0.5
1M
1x
21
Qwen3.8 Flash
Qwen
$0.000385
$0.385
$0.15
$0.47
1M
1x
22
Qwen3.8 Omni Flash
Qwen
$0.000385
$0.385
$0.15
$0.47
1M
1x
23
Ring-2.6-1T
inclusionAI
$0.000388
$0.388
$0.075
$0.625
262K
1x
24
Hy3
Tencent
$0.000396
$0.396
$0.132
$0.528
262K
1x
25
KAT-Coder-Air V2.5
Kwaipilot
$0.000450
$0.450
$0.15
$0.6
256K
1x
26
DeepSeek V3.2
DeepSeek
$0.000469
$0.469
$0.269
$0.4
164K
1x
27
Qwen3 Coder Next
Qwen
$0.000520
$0.520
$0.12
$0.8
262K
1x
28
Llama 4 Maverick
Meta
$0.000600
$0.600
$0.2
$0.8
1M
1x
29
Qwen3.6 35B A3B
Qwen
$0.000640
$0.640
$0.14
$1
262K
1x
30
DeepSeek V3.1
DeepSeek
$0.000725
$0.725
$0.25
$0.95
164K
1x
31
Codestral 2508
Mistral
$0.000750
$0.750
$0.3
$0.9
256K
1x
32
Qwen3.6 Flash
Qwen
$0.000750
$0.750
$0.1875
$1.125
1M
1x
33
Nex-N2-Pro
Nex Agi
$0.000750
$0.750
$0.25
$1
262K
1x
34
MiniMax M2
MiniMax
$0.000765
$0.765
$0.255
$1.02
205K
1x
35
Step 3.7 Flash
Stepfun
$0.000775
$0.775
$0.2
$1.15
262K
1x
36
GPT-5.6 Luna Pro
OpenAI
$0.000800
$0.800
$0.2
$1.2
1M
not metered
37
GPT-5.6 Luna
OpenAI
$0.000800
$0.800
$0.2
$1.2
1M
1x
38
GPT-5.4 Nano
OpenAI
$0.000825
$0.825
$0.2
$1.25
400K
2x
39
MiMo-V2.6-Pro
Xiaomi
$0.000870
$0.870
$0.435
$0.87
1M
2x
40
MiniMax M3
MiniMax
$0.000900
$0.900
$0.3
$1.2
1M
1x
41
LongCat 2.0
Meituan
$0.000900
$0.900
$0.3
$1.2
1M
1x
42
Perceptron Mk1
Perceptron
$0.000900
$0.900
$0.15
$1.5
33K
1x
43
DeepSeek V4.1 Flash
DeepSeek
$0.000900
$0.900
$0.3
$1.2
1M
1x
44
Qwen3.7 Plus
Qwen
$0.000960
$0.960
$0.32
$1.28
1M
1x
45
GLM 5.3 FlashX
Z.ai
$0.000995
$0.995
$0.37
$1.25
1M
2x
46
Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image)
Google
$0.0010
$1.000
$0.25
$1.5
66K
not metered
47
Gemini 3.1 Flash Lite
Google
$0.0010
$1.000
$0.25
$1.5
1M
1x
48
Inkling Small
Thinkingmachines
$0.0010
$1.050
$0.45
$1.2
524K
1x
49
Command A+
Cohere
$0.0010
$1.050
$0.3
$1.5
192K
2x
50
Muse Glimmer 30B
Meta
$0.0011
$1.100
$0.35
$1.5
131K
2x
51
Qwen3.5 Plus 2026-04-20
Qwen
$0.0012
$1.200
$0.3
$1.8
1M
1x
52
Mistral Large 3 2512
Mistral
$0.0013
$1.250
$0.5
$1.5
262K
2x
53
GLM 4.7
Z.ai
$0.0013
$1.275
$0.4
$1.75
205K
2x
54
Gemini 3.7 Flash
Google
$0.0013
$1.313
$0.375
$1.875
1M
2x
55
Mistral Medium 3.1
Mistral
$0.0014
$1.400
$0.4
$2
131K
2x
56
Aion-3.0-Mini
Aion Labs
$0.0014
$1.400
$0.7
$1.4
131K
2x
57
Aion 3.5 Mini
AionLabs
$0.0014
$1.400
$0.7
$1.4
262K
3x
58
GLM 4.6
Z.ai
$0.0015
$1.500
$0.5
$2
205K
1x
59
Gemini 2.5 Flash
Google
$0.0015
$1.550
$0.3
$2.5
1M
2x
60
Gemini 3.5 Flash Lite
Google
$0.0015
$1.550
$0.3
$2.5
1M
2x
61
GLM 5
Z.ai
$0.0016
$1.560
$0.6
$1.92
205K
2x
62
Kimi K2 0711
Moonshot
$0.0017
$1.720
$0.57
$2.3
131K
10x
63
Seed 2.1 Turbo
Bytedance Seed
$0.0018
$1.750
$0.5
$2.5
262K
2x
64
Kimi K2 Thinking
Moonshot
$0.0019
$1.850
$0.6
$2.5
262K
10x
65
R1
DeepSeek
$0.0019
$1.950
$0.7
$2.5
64K
2x
66
Seed-2.0-Code
Bytedance Seed
$0.0020
$2.000
$0.5
$3
262K
3x
67
Nano Banana 2 (Gemini 3.1 Flash Image)
Google
$0.0020
$2.000
$0.5
$3
131K
not metered
68
Grok Build 0.1
xAI
$0.0020
$2.000
$1
$2
256K
3x
69
Qwen3.8 27B
Qwen
$0.0021
$2.050
$0.45
$3.2
1M
3x
70
Hy4 preview
Tencent
$0.0021
$2.084
$0.834
$2.501
1M
3x
71
KAT-Coder-Pro V2.5
Kwaipilot
$0.0022
$2.220
$0.74
$2.96
256K
3x
72
Nova Pro 1.0
Amazon
$0.0024
$2.400
$0.8
$3.2
300K
3x
73
Nemotron 3 Ultra
NVIDIA
$0.0024
$2.400
$0.6
$3.6
512K
3x
74
Kimi K2.7 Code
Moonshot
$0.0025
$2.460
$0.71
$3.5
262K
10x
75
GLM 5.2
Z.ai
$0.0025
$2.484
$0.966
$3.036
1M
3x
76
Grok 4.3
xAI
$0.0025
$2.500
$1.25
$2.5
1M
4x
77
Gemini 3.6 Flash
Google
$0.0026
$2.625
$0.75
$3.75
1M
3x
78
Sakana Namazu
Sakana
$0.0029
$2.950
$0.95
$4
262K
4x
79
DeepSeek V4 Pro 0813
DeepSeek
$0.0030
$2.970
$1.188
$3.564
1M
2x
80
Inkling
Thinkingmachines
$0.0030
$2.975
$0.95
$4.05
1M
4x
81
GPT-5.4 Mini
OpenAI
$0.0030
$3.000
$0.75
$4.5
400K
4x
82
Muse Spark 1.2
Meta
$0.0034
$3.375
$1.25
$4.25
1M
5x
83
Muse Spark 1.1
Meta
$0.0034
$3.375
$1.25
$4.25
1M
5x
84
Muse Spark 1.3
Meta
$0.0034
$3.375
$1.25
$4.25
1M
5x
85
Claude Haiku 4.5
Anthropic
$0.0035
$3.500
$1
$5
200K
4x
86
GLM 5.3
Z.ai
$0.0036
$3.600
$1.4
$4.4
1M
5x
87
Qwen3.7 Max
Qwen
$0.0037
$3.688
$1.475
$4.425
1M
5x
88
Grok 4.7
xAI
$0.0040
$4.000
$1.6
$4.8
500K
6x
89
Grok 4.5
xAI
$0.0050
$5.000
$2
$6
500K
6x
90
Qwen3.8 Max
Qwen
$0.0050
$5.000
$2
$6
1M
6x
91
Qwen3.8 2.4T A95B
Qwen
$0.0050
$5.000
$2
$6
1M
8x
92
Grok 4.6
xAI
$0.0050
$5.000
$2
$6
500K
6x
93
Fugu Max
Sakana
$0.0050
$5.000
$2
$6
1M
8x
94
Mistral Medium 3.5
Mistral
$0.0053
$5.250
$1.5
$7.5
262K
6x
95
Gemini 3.5 Flash
Google
$0.0060
$6.000
$1.5
$9
1M
8x
96
Aion-3.0
Aion Labs
$0.0060
$6.000
$3
$6
131K
10x
97
Aion 3.5
AionLabs
$0.0060
$6.000
$3
$6
262K
10x
98
Claude Sonnet 5
Anthropic
$0.0070
$7.000
$2
$10
1M
10x
99
GPT-6 Sol
OpenAI
$0.0070
$7.000
$2
$10
1M
10x
100
Claude Sonnet 5.5
Anthropic
$0.0070
$7.000
$2
$10
1M
10x
101
GPT-6.1 Sol
OpenAI
$0.0070
$7.000
$2
$10
1M
10x
102
GLM 5.3 Prime
Z.ai
$0.0072
$7.200
$2.8
$8.8
1M
10x
103
Command A
Cohere
$0.0075
$7.500
$2.5
$10
256K
10x
104
Gemini 3.1 Pro Preview
Google
$0.0080
$8.000
$2
$12
1M
10x
105
GPT-5.6 Terra Pro
OpenAI
$0.0080
$8.000
$2
$12
1M
not metered
106
GPT-5.6 Terra
OpenAI
$0.0080
$8.000
$2
$12
1M
4x
107
Nano Banana Pro (Gemini 3 Pro Image)
Google
$0.0080
$8.000
$2
$12
131K
not metered
108
MiMo-V2.6-Pro-UltraSpeed
Xiaomi
$0.0087
$8.700
$4.35
$8.7
1M
15x
109
GPT-5.4
OpenAI
$0.010
$10.000
$2.5
$15
1M
15x
110
GPT-5.6 Sol Pro
OpenAI
$0.010
$10.000
$2.5
$15
1M
not metered
111
GPT-5.6 Sol
OpenAI
$0.010
$10.000
$2.5
$15
1M
20x
112
Qwen3.8 Max Prime
Qwen
$0.010
$10.000
$4
$12
1M
15x
113
Claude Sonnet 4.6
Anthropic
$0.011
$10.500
$3
$15
1M
10x
114
Claude Sonnet 4.5
Anthropic
$0.011
$10.500
$3
$15
1M
10x
115
Kimi K3
Moonshot
$0.011
$10.500
$3
$15
1M
10x
116
Claude Opus 5.5
Anthropic
$0.014
$14.000
$4
$20
1M
20x
117
Claude Opus 5
Anthropic
$0.018
$17.500
$5
$25
1M
20x
118
Claude Opus 4.8
Anthropic
$0.018
$17.500
$5
$25
1M
20x
119
GPT-5.5
OpenAI
$0.020
$20.000
$5
$30
1M
20x
120
GPT Chat Latest
OpenAI
$0.020
$20.000
$5
$30
400K
25x
121
Fugu Ultra
Sakana
$0.020
$20.000
$5
$30
1M
25x
122
Fugu Ultra v2
Sakana
$0.020
$20.000
$5
$30
1M
25x
123
Claude Opus 5 (Fast)
Anthropic
$0.035
$35.000
$10
$50
1M
not metered
124
Claude Fable 5
Anthropic
$0.035
$35.000
$10
$50
1M
100x
125
Claude Opus 4.8 (Fast)
Anthropic
$0.035
$35.000
$10
$50
1M
not metered
126
Claude Opus 4.7 (Fast)
Anthropic
$0.105
$105.000
$30
$150
1M
not metered
Questions
Why price an answer instead of a million tokens?
Because nobody buys a million tokens. Providers quote input and output separately at different rates, so two models can look close per million and differ by a lot on the same real task. Pricing one fixed workload makes the numbers comparable.
Which model is the best value?
Value depends on the task. The useful reading is that the middle of this table is within a few hundredths of a cent per answer, so for most everyday work the price difference between good models is negligible and the choice should be made on output quality.
How often is it updated?
Prices are refreshed from the provider catalogue and the page carries the fetch date, currently August 20, 2026. Rates in this market move often, so check the date before quoting a figure.
What are Whizi credits?
Whizi charges a message against a monthly allowance, and expensive models cost more credits than cheap ones, from 1x up to 20x. The column shows that multiplier so you can see how a provider list rate maps onto a flat subscription.
How to cite this
A dataset is only reused as often as it is easy to credit, so here is the line to copy:
Whizi AI Model Cost Index, prices as of 2026-08-20. https://whizi.io/tools/model-cost-index
Published under CC BY 4.0, which permits commercial reuse as long as the index is named and this page is linked. Writers who need a figure verified before publishing can reach the maintainers at [email protected].