Llama 4 Maverick, introduced on April 5, 2025, is one of the first models in Meta’s Llama 4 family, designed as a natively multimodal model supporting text + image inputs with text outputs. It employs a Mixture-of-Experts (MoE) architecture with 128 experts, activating ~17B parameters per token out of a pool of ~400B total parameters. This design improves scalability, efficiency, and reasoning capacity. Maverick has a 1M-token context window, enabling it to handle large documents, extended conversations, and multimodal reasoning. Its knowledge cutoff is August 2024.
The model is released under the Llama 4 Community License and comes in both base and instruction-tuned (“Instruct”) versions. Maverick is widely deployed via Hugging Face, Google Vertex AI, Amazon Bedrock, and Oracle Cloud, making it one of the most accessible large open-weight models. However, it outputs text only (no image/audio generation) and, while input capacity is huge, output limits are typically much smaller. The MoE design also raises hardware demands, as maintaining 128 experts requires significant compute resources, and Meta’s license introduces restrictions around commercial-scale use.
Drag and drop an image here, or click to browse
Captioning will run automatically
—
Usage
Past 30 DaysLlama 4 Maverick has not yet been evaluated on the current benchmark. The results below are from the legacy version of Vision Evals, our previous benchmark. See the current Vision Evals
| Category | Passed | Score |
|---|---|---|
| Defect Detection | 10 / 15 | 66.7% |
| Document Understanding | 6 / 9 | 66.7% |
| Object Understanding | 9 / 14 | 64.3% |
| Spatial Understanding | 12 / 19 | 63.2% |
| Object Counting | 3 / 10 | 30% |
| Category | Passed | Score |
|---|---|---|
| License Plate Recognition | 28 / 30 | 93.3% |
| Text Recognition | 25 / 30 | 83.3% |
| Focused Scene OCR | 76 / 99 | 76.8% |
| VQA & Extraction | 45 / 60 | 75% |
| Handwritten Math | 6 / 10 | 60% |
Scores based on a single evaluation run · Methodology
View all legacy Vision Evals results →Llama 4 Maverick costs $0.200 per 1M input tokens and $0.696 per 1M output tokens.
Pricing updated Aug 12, 2026
Estimated cost per task vs. Visual Understanding score, for this model and others ranked near it. Upper-left is the sweet spot (high quality, low cost). Based on Vision Evals (legacy) results.
11 of 11 models plotted
| Model | Score | Median tokens | Est. cost / task | Compare |
|---|---|---|---|---|
| Claude Opus 4.8 | 67.2% | 2.2K | $0.012 | Compare |
| Claude Opus 4.7 | 67.2% | 2.6K | $0.015 | Compare |
| Gemma 4 31B | 67.2% | 467 | $0.0001 | Compare |
| Claude Opus 4.6 | 64.2% | 2.3K | $0.014 | Compare |
| GPT-5.4 Nano | 62.7% | 1.8K | $0.0004 | Compare |
| Llama 4 Maverick(this model) | 59.7% | 2.4K | $0.0005 | — |
| Claude Sonnet 4.5 | 59.7% | 2.3K | $0.0092 | Compare |
| Claude Opus 4.1 | 59.7% | 2.1K | $0.040 | Compare |
| Claude Haiku 4.5 | 58.2% | 2.3K | $0.0030 | Compare |
| GPT-5 Nano | 58.2% | 2.7K | $0.0003 | Compare |
| Qwen3.5 397B A17B | 58.2% | 1.5K | $0.0008 | Compare |
Other models worth comparing for similar use cases.
Llama 4 Maverick ships under a custom, model-specific license rather than a standard permissive or restrictive one, so the Llama 4 Maverick license has to be read directly. Custom model licenses range from effectively permissive to research-only.
Uncertainty around licensing can delay or stop a project, and acceptable-use policies attached to custom licenses are binding terms rather than guidance. Review them alongside the Llama 4 Maverick license before production deployment.
If the custom terms rule out your use case, a commercial license from the rights holder is the way through. Roboflow's licensing page lists the supported models whose commercial license is included in a Roboflow plan, so it is worth checking whether Llama 4 Maverick — or a permissively licensed alternative — fits your deployment.
Do not hesitate to reach out with questions for your commercial project — our team will help you start solving business problems on the first call. See Roboflow commercial licensing for the models included in each plan.
Talk to salesThis model is released under a custom license that does not match a standard open-source identifier. Read the full license text linked from the model documentation.
Custom licenses vary widely in what they permit. Many model-specific custom licenses include commercial-use restrictions (e.g., non-commercial weights, named-user limits, or jurisdiction restrictions). Read the full license before deploying commercially.
Custom licenses are model-specific. Always check the per-model License Notes section above and the linked official license text.
License information is provided as a guide and is not legal advice.
Yes. Llama 4 Maverick accepts image input, and on Roboflow's previous vision benchmark it passed 59.7% of visual understanding tasks (#46 of 77) and scored 78.6% on OCR. You can test it on your own image in the demo above.
Llama 4 Maverick has not yet been evaluated on Roboflow's current Vision Evals. The results on this page are from the previous benchmark.
Yes. The demo on this page runs Llama 4 Maverick in the free Roboflow Playground: upload an image and see results in seconds. A free account unlocks unlimited runs.