DataArticle
GLM-5.2 Max reached 1595 on Code Arena: Frontend, surpassing Opus 4.8 and narrowing the gap to Claude Fable 5, while also posting near-parity agentic reliability scores against Opus 4.8 Max.
Z.ai's GLM-5.2 Max posted a leading Code Arena Frontend score of 1595, beating Opus 4.8 and closing in on Claude Fable 5, while also edging out Opus 4.8 Max on an agentic reliability benchmark. ✦ AI generated
AI Twitter Recap (AINews) · Latent Space · 2026-06-26 · original ↗
On frontend coding, Arena reported that GLM-5.2 Max reached 1595 on Code Arena: Frontend, surpassing Opus 4.8 and narrowing the gap to Claude Fable 5. On agentic reliability, PostTrainBench noted 34.29% for GLM 5.2 Max reasoning, narrowly ahead of Opus 4.8 Max at 34.08%, with zero failed runs across 84 runs.
Read full article ↗excerpt · fair-use quotation
Around this claim
This moment responds to
supports → GLM-5.2 is the first open-weight model that feels right as a general coding agent in real harness use, despite minor integration bugs like image inputs bricking a Fireworks API session.Nathan Lambert · Interconnectssupports → Among current frontier open models, Inkling from Thinking Machines is now the strongest American open model with text, image, and audio input, Nemotron 3 Ultra from NVIDIA is a solid pick for long-running agents due to its cheap Mamba-hybrid long-context inference, and GLM-5.2 from Z.ai is currently the best open model for coding.ByteByteGo · ByteByteGo Newslettersupports → GLM-5.2's release has generated a community focal point rivaled only by DeepSeek R1, exceeding even the 'DeepSeek Moment' the author previously used to describe Kimi K2's release.Nathan Lambert · Interconnects