May 1, 2026
Mistral Medium 3.5 lands #32 on the Vals Index
We evaluated Mistral Medium 3.5 on our suite of benchmarks. Here are the key takeaways:
- Mistralโs new reasoning-mode Medium 3.5 lands at #32 of 46 on the Vals Index (52.77%), and #10 of 18 among open-weight models.
- It is a sizable jump over the prior Mistral Large 3 across most of the suite: +28 points on Finance Agent (46.1% vs 18.1%), +29 points on the SWE-bench Verified subset of the Vals Index (64.7% vs 35.3%), +21 points on Terminal-Bench 2.0 (30.3% vs 9.0%), and +13 points on SAGE (37.6% vs 24.6%).
- One regression worth flagging: Case Law (v2) drops 17 points (44.2% vs 61.4%).
All evals were run via the Mistral API at temperature 0.7, top-p 0.95, with reasoning_effort: high and an 80k max output budget.