
Zhipu AI's GLM-5.2 closes in on closed-source leaders in coding marathons
Quick Answer
Zhipu AI's GLM-5.2, an open-source model with a 1-million-token context, closely trails Anthropic's Claude Opus 4.8 by just one percentage point on the FrontierSWE coding benchmark.
Quick Take
However, it still lags behind closed-source models in reasoning tasks, highlighting the competitive landscape in AI coding capabilities.
Key Points
- GLM-5.2 is released under the MIT license, promoting open-source collaboration.
- The model achieves a close performance to Claude Opus 4.8 on coding tasks.
- Despite its strengths, GLM-5.2 still underperforms in reasoning compared to closed-source models.
- The 1-million-token context allows for extensive coding task handling.
- Zhipu AI's advancements signal increasing competition in AI coding solutions.
Source Excerpt
Chinese AI lab Zhipu AI releases GLM-5. 2 with a stable 1-million-token context under the MIT license. On FrontierSWE, a benchmark for hours-long coding tasks, the open-source model trails Anthropic's Claude Opus 4. 8 by just one percentage point. On reasoning, it still falls well behind closed-source rivals.
Want this in your inbox every morning?
Daily brief at your local 8am — bilingual EN/中文, free.
More from The Decoder
See more →
An AI model programmed nonstop for 19 days on a single MirrorCode task that cost $2,600 to run
Epoch AI's MirrorCode benchmark reveals Claude Opus 4.7 as the leader with a 56% solve rate, reconstructing a 16,000-line toolkit in 14 hours. Despite this, all models tested struggle with the most complex tasks, highlighting limitations in current AI capabilities. The single task consumed $2,600 over 19 days, raising questions about cost-effectiveness in AI development.

