o3

Release Date: 4/16/2025

Accuracy (Average)

74.22%

Latency (Average)

49.45s

Avg. Cost (In/Out)

2 / 8

Context Window

200k

Max Output Tokens

100k

Input Modality

Hyperparameter settings
Default Provider : OpenAI

Temperature

Default

Top P

Default

Top K

Default

Max Output Tokens

100,000

Reasoning Effort

high

Benchmarks

Accuracy

Rankings

0.0%

± 0.96
46/ 100

0.0%

± 2.16
15/ 54

0.0%

± 1.87
31/ 54

0.0%

± 0.93
24/ 73

0.0%

± 3.31
27/ 52

0.0%

± 0.85
16/ 107

0.0%

± 1.39
37/ 96

0.0%

± 1.86
25/ 102

0.0%

± 1.03
21/ 108

0.0%

± 0.42
22/ 119

0.0%

± 0.18
6/ 95

0.0%

± 0.34
28/ 101

0.0%

± 0.95
23/ 69
Contact us
Or send us an email at contact@vals.ai
Proprietary Benchmarks (contact us to get access)
Academic Benchmarks

Read about our methodology.