Meta

Llama 4 Maverick Pricing

Llama 4 Maverick is a budget-tier language model from Meta, priced at $0.22 per 1M input tokens and $0.85 per 1M output tokens. 17B active / 400B total (MoE). Prices via hosted providers.

Input
$0.22
per 1M tokens
Output
$0.85
per 1M tokens
Context
1000K
tokens
Max output
32K
tokens

What does Llama 4 Maverick cost in practice?

These are real-world monthly estimates based on common workload sizes. They assume a 2:1 input-to-output ratio (typical of chat + retrieval-augmented apps).

Workload Input / mo Output / mo Monthly cost Annualized
Light prototyping 0.1M tokens 0.1M tokens $0.06 $0.77
Medium production 5.0M tokens 2.5M tokens $3.23 $38.70
Heavy production 50.0M tokens 25.0M tokens $32.25 $387.00

Need to model your own usage? Use the LLM pricing calculator →

Llama 4 Maverick features

Alternatives to Llama 4 Maverick

Cheaper, same-tier alternatives

If Llama 4 Maverick is more than you need, these budget-tier models cost less per token:

Need the cheapest Meta option? Llama 3.3 70B starts at $ 0.12 input / $ 0.30 output per 1M tokens →

FAQ

How much does Llama 4 Maverick cost per 1M tokens?

Llama 4 Maverick costs $0.22 per 1M input tokens and $0.85 per 1M output tokens.

What is Llama 4 Maverick's context window?

Llama 4 Maverick supports a 1,000,000-token context window with a maximum output of 32,000 tokens per response.

What features does Llama 4 Maverick support?

Vision, Tool Use, Streaming.

How much does Llama 4 Maverick cost for medium production usage?

For ~5M input + 2.5M output tokens per month, Llama 4 Maverick costs approximately $3.23/mo (≈$ 38.70/yr).

What are cheaper alternatives to Llama 4 Maverick?

Same-tier alternatives that cost less: Mistral Small 3.1 , Llama 3.3 70B , DeepSeek-V3 .