Skip to content
LLMs
Meta Llama logo

Meta: Llama 4 Maverick

by Meta Llama

Meta: Llama 4 Maverick is a large language model from Meta Llama. It costs $0.200 per million input tokens and $0.696 per million output tokens. Its context window is 1.0M tokens.

Tool callingStructured outputOpen weightsimage input
Input / 1M tokens
$0.200
Output / 1M tokens
$0.696
Cached input / 1M
-

Not supported

Context window
1.0M

tokens

Who serves it cheapest

5 hosts serve Llama 4 Maverick. Same weights, same API - the price difference is pure margin and routing.

Providers serving Meta: Llama 4 Maverick, cheapest first
ProviderInput / 1MOutput / 1MContextThroughputUptime 24h
DigitalOceanCheapest$0.200$0.696128K-99.9%
DeepInfrafp8$0.200$0.8001.0M-99.8%
Novitafp8$0.270$0.8501.0M-99.5%
Parasailfp8$0.350$1.00524K-99.5%
Google$0.350$1.15524K--

The spread between DigitalOcean and Google is 1.7× for identical weights. Quantization and context limits differ, so check both columns before switching.

Benchmarks

Independent scores published alongside the catalogue.

Intelligence index
9.3
Coding index
16.3
Agentic index
0.6

About Llama 4 Maverick

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...

Specifications

Meta: Llama 4 Maverick specifications
Model IDmeta-llama/llama-4-maverick
ProviderMeta Llama
Context window1.0M tokens
Max output115K tokens
Input modalitiestext, image
Output modalitiestext
Knowledge cutoff2024-08-31
Open weightsYes - meta-llama/Llama-4-Maverick-17B-128E-Instruct
ReleasedApril 5, 2025

Cheaper alternatives

Models that cost less than Llama 4 Maverick while keeping at least half its context window and every input modality it supports.

Frequently asked

How much does Meta: Llama 4 Maverick cost?

$0.200 per million input tokens and $0.696 per million output tokens.

What is the context window of Meta: Llama 4 Maverick?

1.0M tokens, with up to 115K tokens of output per request.

Which provider serves Meta: Llama 4 Maverick cheapest?

DigitalOcean at $0.200 per million input tokens - 1.7× cheaper than Google, the most expensive of the 5 hosts serving it.