Meta model product
Llama 4 Maverick is a natively multimodal model capable of processing both text and images.
Updated Aug 12, 2026. Default version: Llama 4 Maverick
Technical details for the model's default version.
This profile uses the model's current scored version. Arena ratings and prices are shown separately.
Benchmark scores for Llama 4 Maverick.
Preference and agent-evaluation results for the default version.
| vision | overall | 99 | 1143.0 | 6,998 | N/A | |
| vision style control | overall | 99 | 1146.8 | 6,998 | N/A | |
| text style control | overall | 223 | 1327.0 | 40,063 | N/A | |
| text | overall | 231 | 1287.6 | 40,063 | N/A |
Provider-specific output speed and catalog latency for Llama 4 Maverick. Runtime does not affect the capability score.
| Sambanova | 638.7 tok/s | 2.04 s | 1M | 1M | |
| Groq | 307.3 tok/s | 0.27 s | 1M | 1M | |
| Together | 97.93 tok/s | 0.2 s | 1M | 1M | |
| Lambda | 93.69 tok/s | 0.65 s | 1M | 1M | |
| DeepInfra | 83.59 tok/s | 0.38 s | 1M | 1M | |
| Novita | 69.42 tok/s | 0.62 s | 1M | 1M | |
| Fireworks | 63.03 tok/s | 0.62 s | 1M | 1M |
Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.
Official vendor API pricing appears first, followed by individual provider offers.
| NanoGPT | meta-llama/llama-4-maverick | global | $0.15 | $0.60 | 1M | |
| Helicone | llama-4-maverick | global | $0.15 | $0.60 | 131.1K | |
| DigitalOcean | llama-4-maverick | global | $0.20 | $0.696 | 128K | |
| OpenRouter | meta-llama/llama-4-maverick | global | $0.20 | $0.696 | 1M | |
| Neon | llama-4-maverick | global | $0.50 | $1.5 | 1M |
Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.
Available versions of this model. The score column identifies the version used in the overall ranking.
| Llama 4 Maverick | 24.9 | 400B | 1M | 1M | No | Llama 4 Community License Agreement |
Key information about Llama 4 Maverick and its available data.
Llama 4 Maverick is a natively multimodal model capable of processing both text and images. It features a 17 billion active parameter mixture-of-experts (MoE) architecture with 128 experts, supporting a wide range of multimodal tasks such as conversational interaction, image analysis, and code generation.
The model includes a 1 million token context window.
Data as of 2026-08-11.
Open a comparison with the three ranked models immediately above and below this model.
Recommendations prioritize the same model type and family, then the closest LLMBoard score.
Common questions about Llama 4 Maverick.
Llama 4 Maverick's default version was released on Apr 5, 2025.
No official standard PAYG price is currently available for Llama 4 Maverick. The lowest tracked third-party offer starts at $0.15 input and $0.60 output via NanoGPT.
Llama 4 Maverick was created by Meta.
The default version has a 1M token context window.
No. The default version is not marked as having publicly available weights.
5 provider offerings are linked to the default version.
Nearby ranked alternatives include Phi 4 Reasoning Plus, DeepSeek-V3, Qwen2 VL 72B.