Release Date: May 19, 2026

Developer Googleย ๐Ÿ‡บ๐Ÿ‡ธ
Context Window 1M
Max Output Tokens 66k
Token Costs (in/out) $1.50/9.00
Weights Private
Input Modalities

Accuracy

53.08 % ยฑ 1.12

Cost / Test (Vals Index)

$ 2.915

Latency

11 min 1 s

Vals Index
BenchmarksAccuracyRankings

0.0%

ยฑ1.12
25/57

0.0%

ยฑ4.10
28/59

0.0%

ยฑ5.79
13/25

0.0%

ยฑ2.75
17/57

0.0%

ยฑ0.23
8/60

0.0%

ยฑ3.21
29/60

0.0%

ยฑ2.11
5/92

0.0%

ยฑ1.92
62/94

0.0%

ยฑ0.91
19/98

0.0%

ยฑ4.63
19/31

0.0%

ยฑ3.44
16/81

0.0%

ยฑ1.28
22/35

0.0%

ยฑ0.85
37/145

0.0%

ยฑ4.73
41/95

0.0%

ยฑ1.46
14/138

0.0%

ยฑ0.95
12/143

0.0%

ยฑ0.85
47/144

0.0%

ยฑ0.31
10/138

0.0%

ยฑ0.77
8/93

0.0%

ยฑ0.00
28/45

0.0%

ยฑ4.47
18/35

0.0%

ยฑ1.83
28/88

0.0%

ยฑ1.12
16/65

0.0%

ยฑ2.44
13/25
Proprietary BenchmarksAcademic BenchmarksIndustry Partners
Vals
Default Provider : Google
Temperature: 1
Top P: Default
Top K: Default
Max Output Tokens: 65,536
Reasoning Effort: high

Updates

May 19, 2026

The model has a 1M-token context window. Evaluations were run using a reasoning effort of โ€œhighโ€, a temperature of 1.0, and max output tokens set to 65k.

Congrats to the Google team on the strong release!