Mar 28, 2025
Gemini 2.5 Pro Exp evaluated on all benchmarks!
We just evaluated Gemini 2.5 Pro Exp on all benchmarks!
- Gemini 2.5 Pro Exp is Googleβs latest experimental model and the new State-of-the-Art, achieving an impressive average accuracy of 82.3% across all benchmarks with a latency of 24.68s.
- The model ranks #1 on many of our benchmarks including CorpFin, Math500, LegalBench, GPQA Diamond, MMLU Pro, and MMMU Pro.
- It excels in academic benchmarks, with standout performances on Math500 (95.2%), MedQA (93.0%), and MGSM (92.2%).
- Gemini 2.5 Pro Exp demonstrates strong legal reasoning capabilities with 86.1% accuracy on CaseLaw and 83.6% on LegalBench, though it scores lower on ContractLaw (64.7%).