Compare Models

Select models (max 5)
Mistral Large 4Gemini 4 Argon
Benchmarks

Vals Index *

Mistral Large 4
0.00%± 1.11
(44/44)
Gemini 4 Argon
0.00%± 0.97
(44/44)

Legal Research Bench *

Mistral Large 4
0.00%± 3.23
(74/74)
Gemini 4 Argon
0.00%± 3.46
(74/74)

Finance Agent (v2) *

Mistral Large 4
0.00%± 0.58
(75/75)
Gemini 4 Argon
0.00%± 0.32
(75/75)

Tax Agent Bench *

Mistral Large 4
0.00%± 3.23
(66/66)
Gemini 4 Argon
0.00%± 2.87
(66/66)

MedCode *

Mistral Large 4
0.00%± 2.17
(105/105)
Gemini 4 Argon
0.00%± 2.10
(105/105)

Terminal-Bench Science

Mistral Large 4
N/A
Gemini 4 Argon
0.00%± 5.98
(38/38)

Code Migration *

Mistral Large 4
0.00%± 4.22
(74/74)
Gemini 4 Argon
0.00%± 4.35
(74/74)

Terminal-Bench 4.0

Mistral Large 4
0.00%± 0.88
(44/44)
Gemini 4 Argon
0.00%± 2.31
(44/44)

Vibe Code Bench v1.1 *

Mistral Large 4
0.00%± 3.55
(109/109)
Gemini 4 Argon
0.00%± 1.90
(109/109)

Overall performance

Performance on the Vals Index, a GDP-weighted aggregation of tasks across finance, coding, and law

2026-10-06

Comparison by Industry

Model performance on different sections of the economy.

Legal
23.78%
54.23%
Finance
58.06%
72.29%
Healthcare
60.56%
73.11%
Math
10.00%
99.00%
Science
39.60%
55.36%
Education
N/A
53.65%
Coding
38.25%
64.03%
Cyber
N/A
61.07%
Social Mobility
N/A
69.76%

Cost Analysis

Per task and per token model pricing.

Vals Index · USD
$13.78
$15.68
Cost / TestVals Index
$13.78
$15.68
Input Cost/ 1M Tokens
$1.36
$4.00
Input Cache Read/ 1M Tokens
$0.14
$0.20
Output Cost/ 1M Tokens
$4.18
$20.00

Model Metadata

Basic information about each model.

Model provider
MistralMistral AI
GoogleGoogle
Latency6350.41s2792.77s
Cost (In/Out)$1.36 / $4.18$4 / $20
Context Window512k1M
Max Output Token256k262k
Input Modality