All benchmarks
Meta AI · USA · Released 2025-04

Llama 4 Maverick

Meta's flagship MoE model — 400B total / 17B active — with native multimodal input and 1M-token context.

Access
open-weights
Context
1,000,000 tokens
Modalities
text, vision
Languages
12 (official)

Benchmark scores

MMLU
85.5%
Massive Multitask Language Understanding
GPQA
69.8%
Graduate-level science QA (Diamond)
HumanEval
88.0%
Python code completion
SWE-bench
N/A
Real GitHub issue resolution (Verified)
LiveCodeBench
N/A
Contamination-free coding
AIME
N/A
American Invitational Mathematics Exam
MMMU
73.4%
Multimodal college-level reasoning
MATH
78.0%
Competition-level math word problems

Scores as reported by the vendor or leading public leaderboards. "N/A" means the score has not been publicly disclosed for this metric.

Explain like I'm 5

A giant open model that sees images and reads huge documents.

Key features
  • Mixture-of-Experts
  • Native multimodal
  • 1M context
  • Open weights
Strengths
  • Open weights
  • Great cost/perf
  • Long context
Limitations
  • Below closed frontier on hardest reasoning

Best for

Self-hosted assistantsFine-tuningOn-prem RAG

Related models