Mar 24, 2025
Command A evaluated on all benchmarks!
We just evaluated Command A on all benchmarks!
- Command A is Cohereβs most efficient and performant model to date, specializing in agentic AI, multilingual, and human evaluations for real-life use cases.
- On our proprietary benchmarks, Command A shows mixed performance, ranking 23rd out of 28 models on TaxEval but a good 10th out of 22 models on CorpFin.
- The model performs better on some academic benchmarks, scoring 78.7% on LegalBench (9th place) and 86.8% on MGSM (13th place).
- However, it struggles with AIME (13.3%, 12th place) and GPQA Diamond (29.3%, 18th place).