# Llama 4 Maverick

> Llama 4 Maverick is an open-weight language model from Meta, released on 5 Apr 2025. It accepts text and images and generates text, with a 1M-token context window and up to 16K output tokens per response. As of 4 Oct 2026, the lowest listed API price is $0.19 per million input tokens and $0.65 per million output tokens (OpenRouter), across 1 provider we track.

Source page: https://www.aimodel.directory/models/llama-4-maverick
Last verified: 4 Oct 2026

## Specifications

- **Developer:** Meta
- **Release date:** 5 Apr 2025 (source: OpenRouter, https://openrouter.ai/meta-llama/llama-4-maverick)
- **Status:** Generally available
- **Weights:** Open weights (source: Hugging Face, https://huggingface.co/meta-llama/Llama-4-Maverick-17B-128E-Instruct)
- **License:** llama4 (source: Hugging Face, https://huggingface.co/meta-llama/Llama-4-Maverick-17B-128E-Instruct)
- **Parameters:** 402B (source: Hugging Face, https://huggingface.co/meta-llama/Llama-4-Maverick-17B-128E-Instruct)
- **Context window:** 1M (1,048,576 tokens) (source: OpenRouter, https://openrouter.ai/meta-llama/llama-4-maverick)
- **Max output:** 16K tokens (source: OpenRouter, https://openrouter.ai/meta-llama/llama-4-maverick)
- **Knowledge cutoff:** 31 Aug 2024 (source: OpenRouter, https://openrouter.ai/meta-llama/llama-4-maverick)
- **Input:** Text and images (source: OpenRouter, https://openrouter.ai/meta-llama/llama-4-maverick)
- **Output:** Text (source: OpenRouter, https://openrouter.ai/meta-llama/llama-4-maverick)
- **Reasoning mode:** No (source: OpenRouter, https://openrouter.ai/meta-llama/llama-4-maverick)
- **Tool calling:** Yes (source: OpenRouter, https://openrouter.ai/meta-llama/llama-4-maverick)
- **Structured output:** Yes (source: OpenRouter, https://openrouter.ai/meta-llama/llama-4-maverick)
- **Hugging Face:** meta-llama/Llama-4-Maverick-17B-128E-Instruct (source: Hugging Face, https://huggingface.co/meta-llama/Llama-4-Maverick-17B-128E-Instruct)

## API pricing (USD per 1M tokens)

| Provider | Input | Output | Cached input | Context |
|---|---|---|---|---|
| OpenRouter | $0.19 | $0.65 | – | 128K |

## FAQ

### What is the context window of Llama 4 Maverick?

Llama 4 Maverick has a context window of 1,048,576 tokens (1M) and can generate up to 16,384 tokens in a single response, according to OpenRouter.

### How much does the Llama 4 Maverick API cost?

Llama 4 Maverick is listed from $0.19 per million input tokens and $0.65 per million output tokens (OpenRouter). Prices last verified 4 Oct 2026.

### Is Llama 4 Maverick open source?

Llama 4 Maverick is an open-weight model: its weights are published on Hugging Face as meta-llama/Llama-4-Maverick-17B-128E-Instruct under the llama4 license. Check the license terms before commercial use.

### When was Llama 4 Maverick released?

Meta released Llama 4 Maverick on 5 Apr 2025, according to OpenRouter.

---
Data licensed CC BY 4.0. Cite as: AI Model Directory, "Llama 4 Maverick", https://www.aimodel.directory/models/llama-4-maverick