All benchmarks
Cohere · Canada · Released 2025-03

Command A

Cohere's enterprise-grade model tuned for RAG, agents and multilingual business workflows.

Access
open-weights
Context
256,000 tokens
Modalities
text
Languages
23

Benchmark scores

MMLU
85.0%
Massive Multitask Language Understanding
GPQA
N/A
Graduate-level science QA (Diamond)
HumanEval
86.0%
Python code completion
SWE-bench
N/A
Real GitHub issue resolution (Verified)
LiveCodeBench
N/A
Contamination-free coding
AIME
N/A
American Invitational Mathematics Exam
MMMU
N/A
Multimodal college-level reasoning
MATH
N/A
Competition-level math word problems

Scores as reported by the vendor or leading public leaderboards. "N/A" means the score has not been publicly disclosed for this metric.

Explain like I'm 5

Built for companies — great at searching their documents and using tools.

Key features
  • RAG-optimized
  • Tool use
  • Multilingual
Strengths
  • Enterprise RAG
  • Groundedness
Limitations
  • Not aimed at pure reasoning benchmarks

Best for

Enterprise searchGrounded chat

Related models