GLM-5.2 is the step change for open agents
| Source: Interconnects (Nathan Lambert)
Tags: GLM-5.2, Z.ai, open-source models, SLIME, agent benchmarks, MIT license
Z.ai's MIT-licensed GLM-5.2 has become the first open-weight model to genuinely compete with OpenAI and Anthropic's top closed models on agent benchmarks, matching Opus 4.8 on Arena's agent leaderboard in max-thinking mode and outperforming Gemini across multiple evals.
Details
Z.ai released GLM-5.2 on June 13, 2026 — first to their Coding Plan members, then as MIT-licensed public weights on June 16. The timing was deliberate: the release coincided with backlash against Anthropic's export restrictions on Claude Fable 5, a marketing opportunity Chinese open-weight labs have consistently exploited. The model's significance is not in benchmark numbers alone. What stands out is GLM-5.2's position on Arena's agent leaderboard as the only open-weight model competing with OpenAI and Anthropic's flagship models. It reportedly matches Opus 4.8 at no-thinking effort when running at max thinking mode. It also topped the Design Arena benchmark, outperforming Claude Fable directly. Nathan Lambert (Interconnects) frames this as a capability threshold — the point where an open model's user experience meaningfully improves across new use cases. Z.ai trained GLM-5.2 using their SLIME RL framework; the recommendation is to always run the model at max thinking effort. Community benchmarks following the release corroborated the official scores across multiple evaluation frameworks. The broader context: Moonshot AI (Kimi) and Z.ai (GLM) now dominate the reputation market for trusted open-weight models among AI researchers. GLM-5.2 represents the first convincing entry of an open model into genuine competition with closed-source frontier models on agent tasks — the category where the gap has been most persistent.