Gemini 3.1 Pro Preview (02/26)

Gemini 3.1 Pro Preview (02/26) is a model from Google, released February 19, 2026. It ranks #39 of 43 models on the Vals Index with 33.44%. Its best result is #2 of 104 on MedCode.

Release Date: Feb 19, 2026

Developer Google Β πŸ‡ΊπŸ‡Έ
Context Window 1M
Max Output Tokens 66k
Token Costs (in/out) $2.00/12.00
Weights Private
Input Modalities

Accuracy

33.44 % Β± 1.13

Cost / Test (Vals Index)

$ 2.180

Latency

9 min 52 s

Vals Index
BenchmarksAccuracyRankings

33.44%

Β±1.13
39/43

17.31%

Β±3.93
49/73

52.62%

Β±2.97
48/70

42.98%

Β±1.21
53/74

20.67%

Β±2.81
50/73

59.06%

Β±2.00
2/104

76.11%

Β±1.92
74/106

69.40%

Β±0.91
6/98

20.72%

Β±2.73
19/23

26.00%

Β±4.41
35/46

48.68%

Β±3.29
26/90

53.79%

Β±1.30
40/47

37.84%

Β±3.10
50/65

72.88%

Β±0.86
58/145

6.69%

Β±2.40
20/20

32.03%

Β±4.34
65/108

60.00%

Β±2.94
22/23

95.45%

Β±1.05
1/138

51.83%

Β±2.59
22/40

88.48%

Β±0.93
6/143

87.40%

Β±0.33
4/149

90.99%

Β±0.28
4/138

88.21%

Β±0.78
10/93

0.00%

Β±0.00
46/62

78.80%

Β±1.83
29/88

2.52%

Β±0.51
37/43

1.43%

Β±1.43
32/38
Proprietary BenchmarksAcademic BenchmarksIndustry Partners
Vals
Default Provider : Google
Temperature: 1
Top P: Default
Top K: Default
Max Output Tokens: 65,536
Reasoning Effort: high

Compared with

Updates

Feb 20, 2026

We evaluated Gemini 3.1 Pro Preview (02/26) across our full benchmark suite. Here are the key takeaways:

One notable metric to call out here is the model achieves this performance at a lower cost than models like Claude Opus 4.6, Claude Sonnet 4.6, GPT 5.2 and O3.

Evaluations were run with a temperature of 1.0 and a β€œhigh” thinking level, via the official Google API.

Congratulations to the Google team on another outstanding model!

Feb 19, 2026

We evaluated Gemini 3.1 Pro Preview (02/26) on our Vals Index. Here are the key takeaways:

We are evaluating Gemini 3.1 Pro Preview (02/26) on our full suite of benchmarks and will be sharing updates soon!