All benchmarks
Mistral AI · France · Released 2024-11

Mistral Large 2

Europe's flagship dense 123B model — strong multilingual + code performance with open weights for research.

Access
open-weights
Context
128,000 tokens
Modalities
text
Languages
80+

Benchmark scores

MMLU
84.0%
Massive Multitask Language Understanding
GPQA
N/A
Graduate-level science QA (Diamond)
HumanEval
92.0%
Python code completion
SWE-bench
N/A
Real GitHub issue resolution (Verified)
LiveCodeBench
N/A
Contamination-free coding
AIME
N/A
American Invitational Mathematics Exam
MMMU
N/A
Multimodal college-level reasoning
MATH
71.5%
Competition-level math word problems

Scores as reported by the vendor or leading public leaderboards. "N/A" means the score has not been publicly disclosed for this metric.

Explain like I'm 5

A European model that speaks lots of languages and writes solid code.

Key features
  • Function calling
  • Multilingual
  • Open weights
Strengths
  • Multilingual
  • European data residency
Limitations
  • No native multimodal

Best for

EU deploymentsMultilingual apps

Related models