The same answer costs 2,000 times more on one model than another.
Providers publish price per million tokens, which is not comparable across models. This index prices one standard answer, 1,000 input tokens plus 500 output tokens, across 100 models from 29 providers. Last updated August 20, 2026.
Headline figures
Cheapest answer: $0.000053 on Ling-3.0-flash (inclusionAI).
Median answer: $0.0019 on Kimi K2 Thinking.
Most expensive answer: $0.105 on Claude Opus 4.7 (Fast) (Anthropic).
Spread from cheapest to most expensive: about 2,000 times.
How this is calculated
One standard answer is 1,000 input tokens, roughly a prompt with a couple of paragraphs of context, and 500 output tokens, roughly a substantial reply. Every model is priced on that identical workload: (input rate x 1,000 + output rate x 500) divided by 1,000,000.
Change the mix and the ordering shifts, which is the point: a model with cheap input and dear output wins on short prompts and loses on long replies. The formula is published so you can recompute it against your own mix.
Sources and limits
Prices are public list rates from the OpenRouter model catalogue, in USD per million tokens, fetched August 20, 2026.
They exclude prompt caching discounts, batch pricing, and negotiated volume rates. Real bills at scale run lower.
Cost is not quality. A model 2,000 times cheaper is not 2,000 times worse, and on many everyday tasks the gap is small.
The credits column is what Whizi charges per message on its own plans, shown for the 89 models it meters individually. It is not a provider rate.
The full index
100 models, cheapest first. Free to reuse with attribution to the Whizi AI Model Cost Index and a link to this page. Raw data: JSON and CSV.
#
Model
Provider
Per answer
Per 1,000 answers
$ / 1M in
$ / 1M out
Context
Whizi credits
1
Ling-3.0-flash
inclusionAI
$0.000053
$0.052
$0.021
$0.063
262K
1x
2
Nex-N2-Mini
Nex Agi
$0.000075
$0.075
$0.025
$0.1
262K
1x
3
Solar Pro 4
Upstage
$0.000090
$0.090
$0.03
$0.12
524K
1x
4
Qwen3.7 Flash
Qwen
$0.000095
$0.095
$0.03
$0.13
1M
1x
5
Granite 4.1 8B
Ibm Granite
$0.000100
$0.100
$0.05
$0.1
131K
1x
6
Laguna XS 2.1
Poolside
$0.000120
$0.120
$0.06
$0.12
262K
1x
7
Hy-MT2-1.8B
Tencent
$0.000133
$0.133
$0.044
$0.177
8K
not metered
8
Phi 4
Microsoft
$0.000140
$0.140
$0.07
$0.14
16K
1x
9
Nemotron 3.5 Lightning
NVIDIA
$0.000180
$0.180
$0.08
$0.2
262K
1x
10
Laguna S 2.1
Poolside
$0.000180
$0.180
$0.09
$0.18
1M
1x
11
Hy-MT2-30B-A3B
Tencent
$0.000222
$0.222
$0.074
$0.295
8K
not metered
12
Llama 3.3 70B Instruct
Meta
$0.000260
$0.260
$0.1
$0.32
131K
1x
13
DeepSeek V4 Flash 0731
DeepSeek
$0.000280
$0.280
$0.14
$0.28
1M
1x
14
Gemini 2.5 Flash Lite
Google
$0.000300
$0.300
$0.1
$0.4
1M
1x
15
Ring-2.6-1T
inclusionAI
$0.000388
$0.388
$0.075
$0.625
262K
1x
16
Hy3
Tencent
$0.000396
$0.396
$0.132
$0.528
262K
1x
17
KAT-Coder-Air V2.5
Kwaipilot
$0.000450
$0.450
$0.15
$0.6
256K
1x
18
DeepSeek V3.2
DeepSeek
$0.000469
$0.469
$0.269
$0.4
164K
1x
19
Qwen3 Coder Next
Qwen
$0.000520
$0.520
$0.12
$0.8
262K
1x
20
Llama 4 Maverick
Meta
$0.000600
$0.600
$0.2
$0.8
1M
1x
21
Qwen3.6 35B A3B
Qwen
$0.000640
$0.640
$0.14
$1
262K
1x
22
DeepSeek V3.1
DeepSeek
$0.000725
$0.725
$0.25
$0.95
164K
1x
23
Codestral 2508
Mistral
$0.000750
$0.750
$0.3
$0.9
256K
1x
24
Qwen3.6 Flash
Qwen
$0.000750
$0.750
$0.1875
$1.125
1M
1x
25
Nex-N2-Pro
Nex Agi
$0.000750
$0.750
$0.25
$1
262K
1x
26
MiniMax M2
MiniMax
$0.000765
$0.765
$0.255
$1.02
205K
1x
27
Step 3.7 Flash
Stepfun
$0.000775
$0.775
$0.2
$1.15
262K
1x
28
GPT-5.6 Luna Pro
OpenAI
$0.000800
$0.800
$0.2
$1.2
1M
not metered
29
GPT-5.6 Luna
OpenAI
$0.000800
$0.800
$0.2
$1.2
1M
1x
30
GPT-5.4 Nano
OpenAI
$0.000825
$0.825
$0.2
$1.25
400K
2x
31
MiniMax M3
MiniMax
$0.000900
$0.900
$0.3
$1.2
1M
1x
32
LongCat 2.0
Meituan
$0.000900
$0.900
$0.3
$1.2
1M
1x
33
Perceptron Mk1
Perceptron
$0.000900
$0.900
$0.15
$1.5
33K
1x
34
Qwen3.7 Plus
Qwen
$0.000960
$0.960
$0.32
$1.28
1M
1x
35
Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image)
Google
$0.0010
$1.000
$0.25
$1.5
66K
not metered
36
Gemini 3.1 Flash Lite
Google
$0.0010
$1.000
$0.25
$1.5
1M
1x
37
Inkling Small
Thinkingmachines
$0.0010
$1.050
$0.45
$1.2
524K
1x
38
Muse Glimmer 30B
Meta
$0.0011
$1.100
$0.35
$1.5
131K
2x
39
Qwen3.5 Plus 2026-04-20
Qwen
$0.0012
$1.200
$0.3
$1.8
1M
1x
40
Mistral Large 3 2512
Mistral
$0.0013
$1.250
$0.5
$1.5
262K
2x
41
GLM 4.7
Z.ai
$0.0013
$1.275
$0.4
$1.75
205K
2x
42
Gemini 3.7 Flash
Google
$0.0013
$1.313
$0.375
$1.875
1M
2x
43
Mistral Medium 3.1
Mistral
$0.0014
$1.400
$0.4
$2
131K
2x
44
Aion-3.0-Mini
Aion Labs
$0.0014
$1.400
$0.7
$1.4
131K
2x
45
GLM 4.6
Z.ai
$0.0015
$1.500
$0.5
$2
205K
1x
46
Gemini 2.5 Flash
Google
$0.0015
$1.550
$0.3
$2.5
1M
2x
47
Gemini 3.5 Flash Lite
Google
$0.0015
$1.550
$0.3
$2.5
1M
2x
48
GLM 5
Z.ai
$0.0016
$1.560
$0.6
$1.92
205K
2x
49
Kimi K2 0711
Moonshot
$0.0017
$1.720
$0.57
$2.3
131K
10x
50
Seed 2.1 Turbo
Bytedance Seed
$0.0018
$1.750
$0.5
$2.5
262K
2x
51
Kimi K2 Thinking
Moonshot
$0.0019
$1.850
$0.6
$2.5
262K
10x
52
R1
DeepSeek
$0.0019
$1.950
$0.7
$2.5
64K
2x
53
Seed-2.0-Code
Bytedance Seed
$0.0020
$2.000
$0.5
$3
262K
3x
54
Nano Banana 2 (Gemini 3.1 Flash Image)
Google
$0.0020
$2.000
$0.5
$3
131K
not metered
55
Grok Build 0.1
xAI
$0.0020
$2.000
$1
$2
256K
3x
56
Qwen3.8 27B
Qwen
$0.0021
$2.050
$0.45
$3.2
1M
3x
57
KAT-Coder-Pro V2.5
Kwaipilot
$0.0022
$2.220
$0.74
$2.96
256K
3x
58
Nova Pro 1.0
Amazon
$0.0024
$2.400
$0.8
$3.2
300K
3x
59
Nemotron 3 Ultra
NVIDIA
$0.0024
$2.400
$0.6
$3.6
512K
3x
60
Kimi K2.7 Code
Moonshot
$0.0025
$2.460
$0.71
$3.5
262K
10x
61
GLM 5.2
Z.ai
$0.0025
$2.484
$0.966
$3.036
1M
3x
62
Grok 4.3
xAI
$0.0025
$2.500
$1.25
$2.5
1M
4x
63
Gemini 3.6 Flash
Google
$0.0026
$2.625
$0.75
$3.75
1M
3x
64
Sakana Namazu
Sakana
$0.0029
$2.950
$0.95
$4
262K
4x
65
DeepSeek V4 Pro 0813
DeepSeek
$0.0030
$2.970
$1.188
$3.564
1M
2x
66
Inkling
Thinkingmachines
$0.0030
$2.975
$0.95
$4.05
1M
4x
67
GPT-5.4 Mini
OpenAI
$0.0030
$3.000
$0.75
$4.5
400K
4x
68
Muse Spark 1.2
Meta
$0.0034
$3.375
$1.25
$4.25
1M
5x
69
Muse Spark 1.1
Meta
$0.0034
$3.375
$1.25
$4.25
1M
5x
70
Claude Haiku 4.5
Anthropic
$0.0035
$3.500
$1
$5
200K
4x
71
GLM 5.3
Z.ai
$0.0036
$3.600
$1.4
$4.4
1M
5x
72
Qwen3.7 Max
Qwen
$0.0037
$3.688
$1.475
$4.425
1M
5x
73
Grok 4.5
xAI
$0.0050
$5.000
$2
$6
500K
6x
74
Qwen3.8 Max
Qwen
$0.0050
$5.000
$2
$6
1M
6x
75
Qwen3.8 2.4T A95B
Qwen
$0.0050
$5.000
$2
$6
1M
8x
76
Grok 4.6
xAI
$0.0050
$5.000
$2
$6
500K
6x
77
Mistral Medium 3.5
Mistral
$0.0053
$5.250
$1.5
$7.5
262K
6x
78
Gemini 3.5 Flash
Google
$0.0060
$6.000
$1.5
$9
1M
8x
79
Aion-3.0
Aion Labs
$0.0060
$6.000
$3
$6
131K
10x
80
Claude Sonnet 5
Anthropic
$0.0070
$7.000
$2
$10
1M
10x
81
Command A
Cohere
$0.0075
$7.500
$2.5
$10
256K
10x
82
Gemini 3.1 Pro Preview
Google
$0.0080
$8.000
$2
$12
1M
10x
83
GPT-5.6 Terra Pro
OpenAI
$0.0080
$8.000
$2
$12
1M
not metered
84
GPT-5.6 Terra
OpenAI
$0.0080
$8.000
$2
$12
1M
4x
85
Nano Banana Pro (Gemini 3 Pro Image)
Google
$0.0080
$8.000
$2
$12
131K
not metered
86
GPT-5.4
OpenAI
$0.010
$10.000
$2.5
$15
1M
15x
87
GPT-5.6 Sol Pro
OpenAI
$0.010
$10.000
$2.5
$15
1M
not metered
88
GPT-5.6 Sol
OpenAI
$0.010
$10.000
$2.5
$15
1M
20x
89
Claude Sonnet 4.6
Anthropic
$0.011
$10.500
$3
$15
1M
10x
90
Claude Sonnet 4.5
Anthropic
$0.011
$10.500
$3
$15
1M
10x
91
Kimi K3
Moonshot
$0.011
$10.500
$3
$15
1M
10x
92
Claude Opus 5
Anthropic
$0.018
$17.500
$5
$25
1M
20x
93
Claude Opus 4.8
Anthropic
$0.018
$17.500
$5
$25
1M
20x
94
GPT-5.5
OpenAI
$0.020
$20.000
$5
$30
1M
20x
95
GPT Chat Latest
OpenAI
$0.020
$20.000
$5
$30
400K
25x
96
Fugu Ultra
Sakana
$0.020
$20.000
$5
$30
1M
25x
97
Claude Opus 5 (Fast)
Anthropic
$0.035
$35.000
$10
$50
1M
not metered
98
Claude Fable 5
Anthropic
$0.035
$35.000
$10
$50
1M
100x
99
Claude Opus 4.8 (Fast)
Anthropic
$0.035
$35.000
$10
$50
1M
not metered
100
Claude Opus 4.7 (Fast)
Anthropic
$0.105
$105.000
$30
$150
1M
not metered
Questions
Why price an answer instead of a million tokens?
Because nobody buys a million tokens. Providers quote input and output separately at different rates, so two models can look close per million and differ by a lot on the same real task. Pricing one fixed workload makes the numbers comparable.
Which model is the best value?
Value depends on the task. The useful reading is that the middle of this table is within a few hundredths of a cent per answer, so for most everyday work the price difference between good models is negligible and the choice should be made on output quality.
How often is it updated?
Prices are refreshed from the provider catalogue and the page carries the fetch date, currently August 20, 2026. Rates in this market move often, so check the date before quoting a figure.
What are Whizi credits?
Whizi charges a message against a monthly allowance, and expensive models cost more credits than cheap ones, from 1x up to 20x. The column shows that multiplier so you can see how a provider list rate maps onto a flat subscription.
How to cite this
A dataset is only reused as often as it is easy to credit, so here is the line to copy:
Whizi AI Model Cost Index, prices as of 2026-08-20. https://whizi.io/tools/model-cost-index
Published under CC BY 4.0, which permits commercial reuse as long as the index is named and this page is linked. Writers who need a figure verified before publishing can reach the maintainers at [email protected].