Accuracy (Vals Index)
63.95% ± 1.31
Latency (Vals Index)
859.57s
Cost/Test (Vals Index)
$0.06
Context Window
1M
Max Output Tokens
384k
Input Modality
Hyperparameter settings
Default Provider :
DeepSeek
Some benchmarks may use different provider and parameters. Please refer to the benchmark page for more information.
Temperature
1
Top P
Default
Top K
Default
Max Output Tokens
384,000
Reasoning Effort
high
Show rankings only among open weight models
Benchmarks
Accuracy
Rankings
Contact us
Read our methodology.