Release Date: Mar 17, 2025

Developer MistralΒ πŸ‡«πŸ‡·
Context Window 131k
Max Output Tokens 8k
Token Costs (in/out) $0.07/0.30
Weights Private
Input Modalities

Accuracy

55.82 %

Avg. Cost (In/Out)

$ 0.07 / $ 0.30

Latency

6.67 s

Vals Index
BenchmarksAccuracyRankings

0.0%

Β±0.96
59/95

0.0%

Β±0.97
131/141

0.0%

Β±2.47
125/133

0.0%

Β±1.11
133/138

0.0%

Β±0.51
117/137

0.0%

Β±0.46
124/133

0.0%

Β±1.18
80/89
Proprietary BenchmarksAcademic BenchmarksIndustry Partners
Vals
Default Provider : Mistral
Temperature: 0.7
Top P: Default
Top K: Default
Max Output Tokens: 8,192

Updates

Apr 4, 2025

We just evaluated Mistral Small 2503 on all benchmarks!

  • Mistral Small 3.1 is Mistral AI’s latest small model, achieving an average accuracy of 61.4% across all benchmarks with a latency of 6.52s - faster than GPT-4o Mini (9.89s) and Llama 3.3 70B (7.67s).
  • Despite its compact size, Mistral Small outperforms Claude 3.5 Haiku (60.2%) in overall accuracy while offering competitive performance to GPT-4o Mini (62.8%).
  • The model excels on MGSM with 85.4% accuracy, comparable to Claude Haiku (85.9%) but behind Llama 3.3 70B’s impressive 91.3%.
  • Like Claude Haiku, the model struggles with AIME (both 3.5%), well behind GPT-4o Mini (11.5%) and Llama 3.3 70B (16.6%).