Gemini 3.5 Flash

Gemini 3.5 Flash is a model from Google, released May 19, 2026. It ranks #34 of 43 models on the Vals Index with 44.79%. Its best result is #6 of 104 on MedCode.

Release Date: May 19, 2026

Developer Google Β πŸ‡ΊπŸ‡Έ
Context Window 1M
Max Output Tokens 66k
Token Costs (in/out) $1.50/9.00
Weights Private
Input Modalities

Accuracy

44.79 % Β± 1.10

Cost / Test (Vals Index)

$ 3.297

Latency

18 min 17 s

Vals Index
BenchmarksAccuracyRankings

44.79%

Β±1.10
34/43

26.75%

Β±4.10
41/73

63.55%

Β±2.75
25/70

57.86%

Β±0.23
11/74

30.77%

Β±3.21
39/73

55.83%

Β±2.11
6/104

76.57%

Β±1.92
72/106

68.12%

Β±0.91
19/98

31.00%

Β±4.65
32/46

49.88%

Β±3.44
18/90

59.47%

Β±1.28
30/47

53.53%

Β±3.30
40/65

74.37%

Β±0.85
37/145

48.68%

Β±4.73
53/108

92.68%

Β±1.46
14/138

87.60%

Β±0.95
12/143

83.60%

Β±0.85
50/149

89.52%

Β±0.31
10/138

88.27%

Β±0.77
8/93

0.00%

Β±0.00
44/62

52.74%

Β±4.47
18/35

78.80%

Β±1.83
28/88

6.06%

Β±1.51
35/43

5.71%

Β±2.79
18/38
Proprietary BenchmarksAcademic BenchmarksIndustry Partners
Vals
Default Provider : Google
Temperature: 1
Top P: Default
Top K: Default
Max Output Tokens: 65,536
Reasoning Effort: high

Compared with

Updates

May 19, 2026

The model has a 1M-token context window. Evaluations were run using a reasoning effort of β€œhigh”, a temperature of 1.0, and max output tokens set to 65k.

Congrats to the Google team on the strong release!