Release Date: Apr 20, 2026

Developer Moonshot AIΒ πŸ‡¨πŸ‡³
Context Window 262k
Max Output Tokens 256k
Token Costs (in/out) $0.95/4.00
Weights Open
Input Modalities

Accuracy

43.47 % Β± 1.17

Cost / Test (Vals Index)

$ 1.850

Latency

39 min 52 s

Vals Index
BenchmarksAccuracyRankings

0.0%

Β±1.17
24/46

0.0%

Β±4.14
21/49

0.0%

Β±6.18
8/22

0.0%

Β±2.74
18/46

0.0%

Β±0.73
28/49

0.0%

Β±2.54
34/49

0.0%

Β±2.04
50/85

0.0%

Β±1.79
47/84

0.0%

Β±0.93
38/95

0.0%

Β±3.43
14/76

0.0%

Β±0.85
31/140

0.0%

Β±4.91
42/84

0.0%

Β±2.00
29/133

0.0%

Β±0.97
18/138

0.0%

Β±0.45
24/137

0.0%

Β±0.33
23/133

0.0%

Β±0.83
19/89

0.0%

Β±0.00
28/39

0.0%

Β±1.91
37/83

0.0%

Β±2.08
36/54
Proprietary BenchmarksAcademic BenchmarksIndustry Partners
Vals
Default Provider : Moonshot AI
Temperature: 0.6
Top P: 0.95
Top K: Default
Max Output Tokens: 256,000

Updates

Apr 20, 2026

Kimi K2.6 is the new #1 open-weight model on our Vals Index! It places #7 among all models at 63.9% accuracy. Here are the key takeaways:

  • The model is a substantial jump over its predecessor Kimi K2.5 (59.6%), and narrowly edges out GLM 5.1 (63.2%) to take the top open-weight spot.
  • Coding is the model’s strongest domain: it places #4 overall on SWE-bench Verified (74.5%) and #7 on Terminal-Bench 2.0 (57.3%), leading all open-weight models on both.
  • It also performs well on Corp Fin (v2), placing #6 at 68.2%.
  • Its weaker showings are on Finance Agent (#19, 57.8%) and Case Law (v2) (#19, 61.2%), where it trails the top closed-weight models by a larger margin.

We ran all benchmarks via the native Moonshot AI provider with a temperature of 1.