Release Date: Jun 13, 2026

Developer zAIΒ πŸ‡¨πŸ‡³
Context Window 1M
Max Output Tokens 131k
Token Costs (in/out) $1.40/4.40
Weights Open
Input Modalities

Accuracy

53.12 % Β± 1.18

Cost / Test (Vals Index)

$ 4.291

Latency

42 min 7 s

Vals Index
BenchmarksAccuracyRankings

0.0%

Β±1.18
17/46

0.0%

Β±4.14
13/49

0.0%

Β±5.49
6/22

0.0%

Β±2.88
15/46

0.0%

Β±0.88
20/49

0.0%

Β±3.22
19/49

0.0%

Β±2.17
43/85

0.0%

Β±2.00
28/84

0.0%

Β±0.00
5/5

0.0%

Β±0.86
47/140

0.0%

Β±4.79
23/84

0.0%

Β±1.77
42/133

0.0%

Β±1.18
88/138

0.0%

Β±0.45
34/137

0.0%

Β±0.33
36/133

0.0%

Β±0.50
7/39

0.0%

Β±4.39
23/28

0.0%

Β±1.69
14/83

0.0%

Β±0.99
20/54
Proprietary BenchmarksAcademic BenchmarksIndustry Partners
Vals
Default Provider : Zhipu AI
Temperature: 1
Top P: 0.95
Top K: Default
Max Output Tokens: 131,072
Reasoning Effort: max

Updates

Jun 17, 2026

  • z.AI just released their latest open-weights GLM 5.2 reasoning model. It’s the #1 open-weight model on the Vals Index (65.02%, #5 overall), reclaiming the top open-weight spot from MiniMax-M3 β€” a 12.5-point jump over its predecessor, GLM 5.1 (52.45%).

  • It’s strongest on coding: #1 open-weight and #3 overall on SWE-bench Verified (82.80%), and #1 open-weight on both Terminal-Bench 2.1 (67.79%) and Vibe Code Bench (63.96%) β€” the latter a 32-point leap from GLM 5.1.

  • It also leads open-weight models on agentic tasks, ranking #1 open-weight on Finance Agent v2 (49.70%) and #1 open-weight on Code Migration (37.87%).

  • On the index it comes in at $2.08/test β€” pricier than MiniMax-M3 ($1.50) and GLM 5.1 ($0.86), but well below frontier closed models like Claude Fable 5 ($5.16).

Eval settings: temperature=1, top_p=0.95, up to 131k max output tokens, run via the native z.AI API. Context window: 1M tokens.

Congrats to the team at z.AI on the release!