Release Date: Dec 22, 2025

Developer zAIΒ πŸ‡¨πŸ‡³
Context Window 200k
Max Output Tokens 128k
Token Costs (in/out) $0.60/2.20
Weights Open
Input Modalities

Accuracy

61.37 %

Avg. Cost (In/Out)

$ 0.60 / $ 2.20

Latency

12 min 48 s

Vals Index
BenchmarksAccuracyRankings

0.0%

Β±2.00
69/85

0.0%

Β±2.12
77/84

0.0%

Β±0.92
96/140

0.0%

Β±2.01
65/133

0.0%

Β±3.40
39/62

0.0%

Β±0.99
51/138

0.0%

Β±0.47
46/137

0.0%

Β±0.37
73/133

0.0%

Β±2.06
62/83
Proprietary BenchmarksAcademic BenchmarksIndustry Partners
Vals
Default Provider : zAI
Temperature: 1
Top P: 1
Top K: Default
Max Output Tokens: 128,000

Updates

Dec 23, 2025

Evaluations are finished for GLM 4.7.

The model was tested with temperature=1 and default top_p for all benchmarks but SWE-bench Verified and Terminal-Bench, which used temperature=0.7 and top_p=1. Reasoning was enabled for all benchmarks.

Dec 22, 2025

GLM 4.7 debuts at #1 on our open-weight leaderboard.

  • It is #9 on the overall leaderboard, behind models from the likes of OpenAI, Gemini, and Anthropic. It is currently the only open-weight model in the top ten.
  • It is a significant performance improvement over GLM 4.6, with a +9.5% performance increase.
  • Despite being priced the same per-token, we also found that it was cheaper than GLM 4.6, using tokens more efficiently.
  • The model uses a new interleaved thinking mode, which may explain the large bump in performance.

The model was tested with temperature=1 and default top_p for all benchmarks but coding benchmarks, which used temperature=0.7 and top_p=1. Reasoning was enabled for all benchmarks.

Results on full benchmarks will be released soon!