Accuracy (Average)

61.21%

Latency (Average)

9.69s

Avg. Cost (In/Out)

3 / 15

Context Window

200k

Max Output Tokens

8k

Input Modality

Hyperparameter settings
Default Provider : Anthropic

Temperature

1

Top P

Default

Top K

Default

Max Output Tokens

8,192

Benchmarks

Accuracy

Rankings

0.0%

± 0.98
62/ 96

0.0%

± 1.80
23/ 68

0.0%

± 0.90
57/ 103

0.0%

± 0.94
88/ 95

0.0%

± 2.47
75/ 98

0.0%

± 1.14
77/ 102

0.0%

± 0.42
74/ 116

0.0%

± 0.67
64/ 95

0.0%

± 0.40
66/ 96

0.0%

± 1.11
45/ 65
Contact us
Or send us an email at contact@vals.ai
Proprietary Benchmarks (contact us to get access)
Academic Benchmarks

Read about our methodology.