Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...
Specifications
| Provider | Meta |
|---|---|
| Context window | 1,048,576 tokens (1.0M) |
| Input price | $0.2 / 1M tokens |
| Output price | $0.8 / 1M tokens |
| Input modalities | text, image |
| Output modalities | text |
| Model ID | meta-llama/llama-4-maverick |
Pricing is per 1M tokens, indicative pricing via OpenRouter (updated 2026-08-22); a provider's own list price may differ.
Compare Llama 4 Maverick
Frequently asked questions
How much does Llama 4 Maverick cost?
Indicative pricing (via OpenRouter, updated 2026-08-22) is $0.2 per 1M input tokens and $0.8 per 1M output tokens. A provider's own list price may differ. On Vincony, Llama 4 Maverick is available on credit-based pricing alongside 800+ other models.
What is Llama 4 Maverick's context window?
Llama 4 Maverick supports a context window of up to 1,048,576 tokens (1.0M), which is how much text it can consider at once.
Who makes Llama 4 Maverick?
Llama 4 Maverick is built by Meta. You can access it — and compare it against other providers' models — through Vincony's multi-model platform.
How do I use Llama 4 Maverick?
You can use Llama 4 Maverick directly through Meta, or via Vincony, which aggregates 800+ models (including Llama 4 Maverick) into one interface with side-by-side comparison and credit-based pricing.