GLM-5.2 is the leading open weights model on Artificial Analysis' Intelligence Index, scoring 51, only behind Fable 5's 60, Opus 4.8's 56, and GPT-5.5's 55
Z ai's GLM-5.2 is the new leading open weights model on the Artificial Analysis Intelligence Index scoring 51 and it sits on the Pareto frontier of Intelligence vs Cost per Task
Artificial Analysis
Context & Ripple Effects
Z.ai had already positioned the GLM-5 line as an MIT-licensed open-weight alternative for reasoning, coding, and agentic work; GLM-5.1 and then GLM-5.2 extended that release cadence. The latest model adds a 1M context window and targets longer-horizon and agentic coding tasks.
Artificial Analysis now places GLM-5.2 ahead of other open-weight models on its Intelligence Index while also putting it on the intelligence-versus-cost-per-task Pareto frontier. It remains behind several leading closed models, making the result more meaningful as a narrowing of the open/closed gap than as outright leadership.
First-order effects
GLM-5.2 becomes the benchmark leader among open-weight models in this index, strengthening Z.ai's position with developers and organizations that value deployable weights alongside capability.
The Pareto-frontier result gives prospective users a third-party basis to evaluate GLM-5.2 on both task intelligence and cost, rather than relying solely on vendor performance claims.
Second-order effects
Other open-weight model providers face pressure to show competitive results not only on capability but also on serving economics, especially for coding and long-horizon workloads highlighted by Z.ai.
Closed-model providers retain higher index scores in this comparison, but a stronger open-weight option can make deployment control and cost more central in model-selection decisions.
Third-order effects
If independent evaluations continue to show open-weight models approaching frontier closed-model performance, the market is likely to segment more sharply between premium proprietary models and configurable models organizations can run or adapt themselves.
The relevant competitive benchmark is shifting from isolated test wins toward an efficiency frontier: capability delivered at a usable task cost, with context length and agentic reliability becoming differentiators.
The trend: Open-weight models are increasingly competing for production AI workloads on the combined dimensions of frontier-adjacent capability, operating cost, and deployability rather than openness alone.
BREAKING: GLM-5.2 is now 1st on Design Arena. With an Elo of 1360, GLM-5.2 has jumped ahead of the now unavailable Claude Fable 5. And it's open weights. This is an improvement of 4 positions and 27 Elo points to achieve one of the highest Elo scores in our code categories [image…
@jietang @teortaxesTex On benchmarks, yes, but as measured by true usefulness even Q1 would be very impressive. Anthropic has rightly focused on maximizing useful intelligence, which does not show up in benchmarks, but definitely shows up in revenue.
Exciting news: GLM-5.2 (Max) ranks #2 in Code Arena: Frontend, with +29pt over Claude Opus 4.7 (Thinking) and only behind Fable 5! GLM-5.2 is the best open model vs Kimi-K2.6 and Minimax-M3 by a large margin. - #2 React and #4 HTML sub-leaderboards - Ranks as the top model in [im…
For the 100th time: the best Chinese model is open source, and this has been the case for almost every week since R1 1.5 years ago I wonder how doomers who coped that “Alibaba is going closed, the trend is clear” feel now
A standout number in Z ai's GLM-5.2 launch is CritPt, a benchmark of unpublished research-level physics problems where it ties with Claude Opus 4.8 and is well above other open weights models Key takeaways: ➤ @Zai_org 's GLM-5.2 (max reasoning effort) leads open weights by a [ima…
For a long time I've been saying that the gap between open source and closed models is going to widen because of the data gap, hardware gap, and increased restrictions on distillation. I was wrong. https://z.ai/ is on another level. Incredible benchmarks on this
Z ai's GLM-5.2 is the new leading open weights model on the Artificial Analysis Intelligence Index scoring 51 and it sits on the Pareto frontier of Intelligence vs Cost per Task @Zai_org's GLM-5.2 is the same size as GLM-5.1 (744B total / 40B active parameters) but scores 11 [ima…
I think GLM 5.2 points to a 7 months gap currently It's around Opus 4.7-4.8 level, all told (modulo vision which in Opus's case is garbage anyway). Mythos reached Preview status (≥ Opus 4.8, functionally) by early Feb 2026. This means full PRC Mythos ("Fable") by Nov-Dec'26. [ima…
A couple of days ago, I claimed Chinese models lag US models by ~7 months. Just spent time with GLM 5.2 (new open model from Zhipu). It has a 1-million token context, excels at long coding projects and agent tasks, and beats GPT-5.5 on several tough benchmarks while matching
I was wrong I've been saying for months that open source AI models are 6 months behind frontier They caught up. GLM 5.2 is as good as Opus 4.8 This changes everything. If you run GLM 5.2 locally no government can take it away. You become sovereign And even if you run through
Just caught up with the recent GLM-5.2 release. The best open-weight model today. Architecture-wise, it's build on the GLM-5 and GLM-5.1 architecture that I covered previously, which means it's reusing the Multi-head Latent Attention (MLA) and DeepSeek Sparse Attention (DSA) [ima…