GLM-5.3: How Chinese labs keep stride with the frontier
Interconnects Nathan Lambert ● Covered by 2 sources
Z.ai says GLM-5.3 is its strongest model yet, and it’s only in the coding plan for now. The surprise: it’s smaller than rivals but already beating some of the frontier on coding tests.
Based on reporting by Interconnects, Nathan Lambert — read the original for the full story.
Summary, retelling and take written by AI under human oversight; images are AI-generated illustrations. How we work · Report an error
Z.ai has pushed out GLM-5.3, and the first version is aimed at coding users. It’s in the company’s coding plan now, with API access promised soon and Hugging Face weights due in two weeks. On the benchmark sheet, it looks unusually strong: Interconnects says it has moved past Moonshot AI’s Kimi K3 on many tests, and even topped Claude Fable 5 or GPT-5.6-Sol on some of them.
The eye-catching part is how little Z.ai claims to have changed. The company’s own line is blunt: “Scaling post-training is all we did for GLM-5.3.” GLM-5.3 uses the same base model as GLM-5.2, but with much heavier post-training. That seems to fit Z.ai’s apparent strength: not necessarily pretraining swagger, but a knack for getting a lot out of the model after the base run is done.
That matters because the model is not huge by the standards of the frontier race. The write-up puts it at roughly 750 billion parameters, which it calls about a third of Kimi K3. And yet it is sitting near the top of agentic coding benchmarks. The release also lands inside a longer Z.ai story: the company has been working on this line for years, from the original GLM work at Tsinghua’s THUDM group through ChatGLM, GLM-4, and now GLM-5.
The broader argument here is less about one model and more about timing. The article says Chinese labs can look faster on public leaderboards because they release sooner, while American labs may spend months in pre-release testing. That extra time gives the Chinese side more room to keep pushing on benchmarks before the next model arrives. If model self-improvement loops start depending on user data, that release speed could matter even more.
Z.ai is also treating this as a security-sensitive release. It says GLM-5.3 is its most capable model yet for cybersecurity work, especially vulnerability discovery, exploit analysis, and multistep security tasks. Broader access comes later, after staged evaluation, and the company says it is watching inference with a request classifier and chain-of-thought monitoring. The company knows what it has built. The rest of the industry seems to know it too.
My take — AI-written commentary, not fact-checked reporting
The real story is not that Z.ai has a miracle model; it’s that the frontier now rewards whoever can train hard, ship faster, and keep the scoreboard happy. That is a lovely setup if one enjoys benchmaxxing as a business strategy and a terrible one if one has to secure actual software. Open weights will make this mess spread faster, because of course they will.
Read more about this at: Interconnects
Related stories
Z.ai Ships GLM-5.3 Without Retraining the Base Model: Better at Complex Coding and Long-Horizon Tasks
MarkTechPost · 3 weeks ago ·
36
GLM-5.3’s Exploits, AI Models and Hardware Speed Up, DeepSeek’s New Agent Harness
The Batch ·
41
GLM-5.2 is the step change for open agents
Interconnects · 2 months ago ·
51