Release Date: Nov 18, 2025

Developer GoogleΒ πŸ‡ΊπŸ‡Έ
Context Window 1M
Max Output Tokens 66k
Token Costs (in/out) $2.00/12.00
Weights Private
Input Modalities

Accuracy

69.38 %

Avg. Cost (In/Out)

$ 2.00 / $ 12.00

Latency

14 min 59 s

Vals Index
BenchmarksAccuracyRankings

0.0%

Β±2.07
12/92

0.0%

Β±1.90
77/94

0.0%

Β±0.91
7/98

0.0%

Β±3.38
28/81

0.0%

Β±0.87
62/145

0.0%

Β±3.06
72/95

0.0%

Β±1.39
19/138

0.0%

Β±0.98
23/143

0.0%

Β±0.37
5/144

0.0%

Β±0.29
7/138

0.0%

Β±0.80
14/93

0.0%

Β±1.90
39/88

0.0%

Β±5.30
15/67
Proprietary BenchmarksAcademic BenchmarksIndustry Partners
Vals
Default Provider : Google
Temperature: 1
Top P: Default
Top K: Default
Max Output Tokens: 65,536
Reasoning Effort: high

Updates

Nov 18, 2025

We just evaluated Gemini 3 Pro (11/25) on our Vals Index! Key takeaways:

Gemini 3 Pro (11/25) demonstrates strong performance across the board, placing first on SAGE, GPQA Diamond and MortgageTax and second on our Multimodal Vals Index.

Gemini 3 Pro (11/25) is three times faster than GPT 5.1, and is also cheaper than Claude Sonnet 4.5 (Thinking), the leader on the Vals Multimodal Index.

Overall, Gemini 3 Pro (11/25) excels in multimodal use cases and provides meaningful improvement over its predecessor, Gemini 2.5 Pro Exp.