Jul 21, 2026
Google's Gemini 3.5 Flash Lite evaluated on the Vals Index
We evaluated Googleโs Gemini 3.5 Flash Lite on the Vals Index and across our benchmark suite.
-
Gemini 3.5 Flash Lite scores 36.71% on the Vals Index, placing #29 of 43 models overall.
-
Its Index component results include 68.63% on the SWE-bench Verified subset, 62.24% on the CorpFin v2 subset, and 46.63% on the Finance Agent v2 subset.
-
The model scores 32.77% on the Vibe Code Bench Index subset and 50.19% across three full trials of Terminal-Bench 2.1.
We evaluated Gemini 3.5 Flash Lite with temperature 1, high reasoning effort, and up to 65k output tokens. The model supports a 1M-token context window, multimodal inputs, and tool calling.
Results are now available across 25 benchmarks. On the full 500-task SWE-bench Verified evaluation, Gemini 3.5 Flash Lite scores 75.00%; the Vals Index uses the 68.63% subset result above.