Release Date: Apr 23, 2026

Developer DeepSeek 🇨🇳
Context Window 1M
Max Output Tokens 384k
Token Costs (in/out) $1.32/3.96
Weights Open
Input Modalities

Accuracy

42.89 % ± 1.19

Cost / Test (Vals Index)

$ 0.898

Latency

23 min 47 s

Vals Index
BenchmarksAccuracyRankings

0.0%

±1.19
34/57

0.0%

±4.04
29/59

0.0%

±5.84
15/25

0.0%

±3.05
37/57

0.0%

±0.65
39/60

0.0%

±2.93
37/60

0.0%

±2.12
52/92

0.0%

±2.00
67/94

0.0%

±3.67
25/31

0.0%

±0.88
69/145

0.0%

±4.77
39/95

0.0%

±1.65
30/138

0.0%

±0.95
14/143

0.0%

±0.47
84/144

0.0%

±0.34
32/138

0.0%

±0.00
36/45

0.0%

±4.60
21/35

0.0%

±1.87
36/88

0.0%

±1.50
52/65
Proprietary BenchmarksAcademic BenchmarksIndustry Partners
Vals
Default Provider : DeepSeek
Temperature: 1
Top P: Default
Top K: Default
Max Output Tokens: 384,000
Reasoning Effort: max

Updates

Apr 24, 2026

DeepSeek is back — DeepSeek V4 just landed #2 open-weight on the Vals Index, narrowly trailing Kimi K2.6 by 0.07%. All rankings below are among open-weight models.

The model has a 1M-token context window, and was run with temp=1, top_p = 0.95, 256k max output tokens, and max reasoning effort via the DeepSeek native API.