Llama 4 Maverick: price, context window and credit cost

The short answer

Llama 4 Maverick is a Meta model priced at $0.2 per million input tokens and $0.8 per million output tokens, with a 1M token context window. One standard answer, meaning 1,000 tokens in and 500 tokens out, costs $0.0006. Inside Whizi it costs 1 credit per message and needs the Pro plan or above.

That makes it one of the cheaper models in the catalogue: 66 of the 89 priced models cost more per answer, and 22 cost less.

What Llama 4 Maverick costs

List rates are what the provider charges per token. The cost of one standard answer is the more useful number, because it prices the same unit of work across every model regardless of how each one is billed.

Input, per million tokens$0.2
Output, per million tokens$0.8
One standard answer$0.0006
One thousand answers$0.6000
Context window1M tokens
Credits per message in Whizi1
Minimum Whizi planPro
Rank by cost, of 89 priced models23

Prices are the published Meta rates as carried by OpenRouter, last refreshed 2026-08-08. The full dataset covering all 100 models from 28 providers is the AI model cost index, which is free to reuse with attribution.

What it costs inside Whizi

Whizi meters messages in credits rather than tokens, and Llama 4 Maverick is charged at 1 credit per message. The cost does not change with the length of your message or the length of the answer, so a one line question and a long document analysis cost the same.

PlanPrice per monthMonthly creditsMessages on this model
Pro$29.99, or $19.99 billed yearly2,0002,000
Powerhouse$49.99, or $34.99 billed yearly8,0008,000

Llama 4 Maverick is included from the Pro plan up. Pro is a curated tier rather than everything cheap: one or two current flagships per model family, each fast tier, and the high volume workhorses. See plans and limits.

How it compares on price

Against the cheapest model in the index, Ling-2.6-flash at $0.000025 per answer, Llama 4 Maverick costs 24 times more. Against the median, GLM 4.6 at $0.0015, it costs 2.5 times less.

Its nearest neighbours from Meta:

ModelOne answerContextCredits
Llama 3.3 70B Instruct$0.0003131K1
Llama 4 Maverick (this page)$0.00061M1
Muse Spark 1.2$0.00341M5

If Llama 4 Maverick is more model than a given task needs, Llama 3.3 70B Instruct is the cheaper Meta option at $0.0003 per answer and 1 credit per message. Routing routine work down a tier is the single biggest thing you can do to make an allowance last.

The context window in practice

Llama 4 Maverick accepts 1M tokens of context, which is roughly 1600 pages of text. That is well above the median for the priced catalogue.

Advertised context and usable context are not the same thing. Recall tends to degrade before the stated limit is reached, particularly for material in the middle of a long input, so the practical test is to upload a long document and ask about something buried halfway through rather than trusting the number.

The number matters most when you are working from your own files. If a document is too long for the model you picked, switching to a larger-context model in the same conversation carries the thread across, so you are not starting over. See switching models mid-conversation.

Workflow checklist
  • Input $0.2 and output $0.8 per million tokens
  • $0.0006 for one standard answer, $0.6000 for a thousand
  • 1M token context window
  • 1 credit per message inside Whizi
  • Requires the Pro plan or above
  • Ranks 23 of 89 priced models by cost per answer
Common questions

Frequently asked questions

How much does Llama 4 Maverick cost?

$0.2 per million input tokens and $0.8 per million output tokens at list rates, which works out to $0.0006 for one standard answer of 1,000 tokens in and 500 out, or $0.6000 for a thousand of them. Inside Whizi it is 1 credit per message on a monthly allowance, rather than metered per token.

Which Whizi plan includes Llama 4 Maverick?

Pro and above, which is $29.99 per month or $19.99 billed yearly. At 1 credits per message, a Pro allowance of 2,000 credits covers 2,000 messages on this model if you spend the whole allowance here, and most people mix in a 1 credit model for routine work so it goes further.

What is the context window of Llama 4 Maverick?

1M tokens, which is roughly 1600 pages of text and sits well above the median for the priced catalogue. Treat that as the ceiling rather than the working figure: recall usually degrades before the stated limit, especially for content in the middle of a long input.

Is Llama 4 Maverick expensive compared to other models?

It ranks 23 of 89 priced models by cost per answer, so 66 cost more and 22 cost less. The spread across the whole catalogue is about 4800 times from cheapest to priciest, which is far wider than most people assume, so "expensive" only means anything relative to the specific alternative you would otherwise use.