DeepSeek V4 Flash 0731: price, context window and credit cost

The short answer

DeepSeek V4 Flash 0731 is a DeepSeek model priced at $0.09 per million input tokens and $0.18 per million output tokens, with a 1M token context window. One standard answer, meaning 1,000 tokens in and 500 tokens out, costs $0.0002. Inside Whizi it costs 1 credit per message and needs the Powerhouse plan or above.

That makes it one of the cheaper models in the catalogue: 80 of the 89 priced models cost more per answer, and 8 cost less.

What DeepSeek V4 Flash 0731 costs

List rates are what the provider charges per token. The cost of one standard answer is the more useful number, because it prices the same unit of work across every model regardless of how each one is billed.

Input, per million tokens$0.09
Output, per million tokens$0.18
One standard answer$0.0002
One thousand answers$0.1800
Context window1M tokens
Credits per message in Whizi1
Minimum Whizi planPowerhouse
Rank by cost, of 89 priced models9

Prices are the published DeepSeek rates as carried by OpenRouter, last refreshed 2026-08-08. The full dataset covering all 100 models from 28 providers is the AI model cost index, which is free to reuse with attribution.

What it costs inside Whizi

Whizi meters messages in credits rather than tokens, and DeepSeek V4 Flash 0731 is charged at 1 credit per message. The cost does not change with the length of your message or the length of the answer, so a one line question and a long document analysis cost the same.

PlanPrice per monthMonthly creditsMessages on this model
Powerhouse$49.99, or $34.99 billed yearly8,0008,000

DeepSeek V4 Flash 0731 is part of the Powerhouse catalogue. Anything not explicitly listed on the Starter or Pro tiers resolves to Powerhouse, which is the deliberate default for the frontier and specialist models. See plans and limits.

How it compares on price

Against the cheapest model in the index, Ling-2.6-flash at $0.000025 per answer, DeepSeek V4 Flash 0731 costs 7.2 times more. Against the median, GLM 4.6 at $0.0015, it costs 8.3 times less.

Its nearest neighbours from DeepSeek:

ModelOne answerContextCredits
DeepSeek V4 Flash 0731 (this page)$0.00021M1
DeepSeek V4 Flash 0423$0.00031M1
DeepSeek V3.2$0.0005164K1
DeepSeek V3.1$0.0007164K1

It is the cheapest DeepSeek model in the priced index, so within this provider there is nothing to route down to.

The context window in practice

DeepSeek V4 Flash 0731 accepts 1M tokens of context, which is roughly 1600 pages of text. That is well above the median for the priced catalogue.

Advertised context and usable context are not the same thing. Recall tends to degrade before the stated limit is reached, particularly for material in the middle of a long input, so the practical test is to upload a long document and ask about something buried halfway through rather than trusting the number.

The number matters most when you are working from your own files. If a document is too long for the model you picked, switching to a larger-context model in the same conversation carries the thread across, so you are not starting over. See switching models mid-conversation.

Workflow checklist
  • Input $0.09 and output $0.18 per million tokens
  • $0.0002 for one standard answer, $0.1800 for a thousand
  • 1M token context window
  • 1 credit per message inside Whizi
  • Requires the Powerhouse plan or above
  • Ranks 9 of 89 priced models by cost per answer
Common questions

Frequently asked questions

How much does DeepSeek V4 Flash 0731 cost?

$0.09 per million input tokens and $0.18 per million output tokens at list rates, which works out to $0.0002 for one standard answer of 1,000 tokens in and 500 out, or $0.1800 for a thousand of them. Inside Whizi it is 1 credit per message on a monthly allowance, rather than metered per token.

Which Whizi plan includes DeepSeek V4 Flash 0731?

Powerhouse and above, which is $49.99 per month or $34.99 billed yearly. At 1 credits per message, a Powerhouse allowance of 8,000 credits covers 8,000 messages on this model if you spend the whole allowance here, and most people mix in a 1 credit model for routine work so it goes further.

What is the context window of DeepSeek V4 Flash 0731?

1M tokens, which is roughly 1600 pages of text and sits well above the median for the priced catalogue. Treat that as the ceiling rather than the working figure: recall usually degrades before the stated limit, especially for content in the middle of a long input.

Is DeepSeek V4 Flash 0731 expensive compared to other models?

It ranks 9 of 89 priced models by cost per answer, so 80 cost more and 8 cost less. The spread across the whole catalogue is about 4800 times from cheapest to priciest, which is far wider than most people assume, so "expensive" only means anything relative to the specific alternative you would otherwise use.