Claude Sonnet 4.5 (Thinking)

Release Date: 9/29/2025

Accuracy (Average)

58.78%

Latency (Average)

1279.97s

Avg. Cost (In/Out)

3 / 15

Context Window

1M

Max Output Tokens

64k

Input Modality

Hyperparameter settings
Default Provider : Anthropic

Temperature

1

Top P

Default

Top K

Default

Max Output Tokens

64,000

Benchmarks

Accuracy

Rankings

0.0%

± 0.96
32/ 109

0.0%

± 2.00
21/ 61

0.0%

± 1.87
12/ 61

0.0%

± 0.95
33/ 78

0.0%

± 3.92
15/ 35

0.0%

± 3.21
38/ 59

0.0%

± 0.86
32/ 115

0.0%

± 3.73
27/ 55

0.0%

± 2.25
36/ 109

0.0%

± 5.92
17/ 54

0.0%

± 1.13
57/ 114

0.0%

± 0.45
21/ 112

0.0%

± 0.39
14/ 108

0.0%

± 0.97
30/ 74

0.0%

± 2.05
31/ 52
Contact us
Or send us an email at contact@vals.ai
Proprietary Benchmarks (contact us to get access)
Academic Benchmarks

Read about our methodology.