All benchmarks
OpenAI · USA · Released 2025-08

GPT-5

OpenAI's flagship reasoning + multimodal model with a unified router that switches between fast and deep-thinking modes.

Access
closed
Context
400,000 tokens
Modalities
text, vision, audio
Languages
100+

Benchmark scores

MMLU
91.4%
Massive Multitask Language Understanding
GPQA
89.4%
Graduate-level science QA (Diamond)
HumanEval
96.3%
Python code completion
SWE-bench
74.9%
Real GitHub issue resolution (Verified)
LiveCodeBench
N/A
Contamination-free coding
AIME
94.6%
American Invitational Mathematics Exam
MMMU
84.2%
Multimodal college-level reasoning
MATH
96.7%
Competition-level math word problems

Scores as reported by the vendor or leading public leaderboards. "N/A" means the score has not been publicly disclosed for this metric.

Explain like I'm 5

The smartest ChatGPT so far. It thinks longer on hard problems and can see, hear and write really well.

Key features
  • Unified reasoning router
  • 400K context
  • Native vision + audio
  • Tool use / agents
  • Long agentic tasks
Strengths
  • State-of-the-art reasoning
  • Strong coding & math
  • Very low hallucination rate vs GPT-4o
Limitations
  • Closed source
  • Rate limits on the highest reasoning tier

Best for

Agentic codingResearch assistantsComplex analysisMultimodal apps

Related models