Skip to main content
AI Choice EngineAI Choice Engine

Meta

Llama 4 Maverick

Llama 4 Maverick is a Meta model with 1M context, Not verified in / Not verified out per million tokens, last verified September 4, 2026. Open weights: yes. Public-board scores are linked to the publisher; vendor-only figures stay in the watch-out, not on the board.

Llama 4 Maverick is Meta's open-weight 400B/17B MoE with a 1M-token context and native multimodal support.

Catalog record checked September 4, 2026Individual provider fields may change

AI model specification details
SpecificationLlama 4 Maverick
ProviderMeta
TierBalanced
Context window1M
Max output131K
Input / 1M tokensNot verifiedUnverified
Output / 1M tokensNot verifiedUnverified
WeightsOpen
Parameters400B total / 17B active (MoE)
Reasoning levelsNot verifiedUnverified
Modalitiestext, image
LicenseLlama 4 Community License
API model idllama-4-maverick
ReleasedApril 5, 2025

Pricing tiers: Open-weight (Llama 4 Community License) — self-host free; hosted inference billed by the serving provider. No first-party Meta API: the Llama API Public Preview was retired 2026-07-06, and Meta's remaining endpoint serves Muse models only.

Verified evidence

Published benchmark results

Each result keeps its source and measurement date visible. A missing benchmark is not treated as a zero.

Published benchmark results for Llama 4 Maverick
BenchmarkScoreMeasuredSource
Artificial Analysis Intelligence Index432026-08-14Artificial Analysis · View source

Best for

  • Open-weight deployments
  • Multimodal understanding
  • Self-hosting

Watch out

Meta has ended Llama development in favour of the proprietary Muse family — Llama stays open weights in maintenance mode. Hosts are dropping it too (Groq retired its Maverick endpoint 2026-03-09); check your host's roadmap before new deployments.

Source receipts

Catalog figures for this model were checked against the following sources.

Common questions

Llama 4 Maverick

Answered from the verified figures on this page rather than general guidance.

What is Llama 4 Maverick's context window?

Llama 4 Maverick accepts about 1M tokens of context. That only matters if you routinely send very long documents, large codebases, or multi-turn histories that approach that limit.

What is Llama 4 Maverick best for?

Llama 4 Maverick is a balanced tier from Meta. It suits open-weight deployments, multimodal understanding, self-hosting. Meta has ended Llama development in favour of the proprietary Muse family — Llama stays open weights in maintenance mode. Hosts are dropping it too (Groq retired its Maverick endpoint 2026-03-09); check your host's roadmap before new deployments.

Can I self-host Llama 4 Maverick?

Llama 4 Maverick publishes open weights, but self-hosting depends on the licence, hardware footprint, quantisation quality, and serving stack. A hosted API is often cheaper until you have measured throughput and concurrency on your own hardware.