China's Z.ai Ships GLM-5.3, Claiming Coding and Cyber Gains Without a Bigger Base Model
Beijing-based Z.ai released GLM-5.3 on August 14, reusing its ~700B-parameter GLM-5.2 base model but claiming a 50% jump on internal coding benchmarks and a leading CyberGym score, positioning it against Anthropic and OpenAI on coding without training a larger model.
Z.ai, the Beijing-based lab formerly known as Zhipu AI, released GLM-5.3 on August 14, the latest in a run of Chinese open-weight models aimed squarely at Anthropic and OpenAI's coding lead.
Same base model, heavier post-training
Unlike most model upgrades, GLM-5.3 doesn't come from a larger or freshly pretrained network. It sits on the same roughly 700-billion-parameter base model as June's GLM-5.2, with Z.ai instead pouring resources into post-training — reinforcement learning and fine-tuning aimed specifically at coding and long-horizon agentic tasks. The company says the approach delivered a 50% improvement over GLM-5.2 on its internal coding-agent benchmark, with the largest gains on the hardest, longest-running tasks: its score on Terminal-Bench 3.0, which tests multi-step terminal and tool-use work, rose from 4.6 to 28.3.
A cybersecurity benchmark stands out
The most notable — and most double-edged — result is on CyberGym, a benchmark that scores a model's ability to find and exploit software vulnerabilities. GLM-5.3 scored 84.5%, more than doubling GLM-5.2's mark and edging out Anthropic's Mythos 5 and OpenAI's GPT-5.6 Sol on the same test, according to Z.ai's own reporting cited by MarkTechPost and Silicon Republic. Stronger exploit-finding ability is useful for legitimate security research and automated patching, but it also raises the familiar dual-use concern that has followed other frontier coding models this year: a model that's better at finding exploits is better for both defenders and attackers.
Availability and what's still missing
GLM-5.3 is live now through Z.ai's API and its GLM Coding Plan, rolled out automatically to existing coding-plan subscribers. The model's weights — which would let developers download and run GLM-5.3 themselves, as they can with GLM-5.2 — are not yet public; Z.ai says it plans to release them within about two weeks, after further safety evaluation given the jump in exploit capability.
The release fits a pattern this year of Chinese labs — Z.ai, DeepSeek, Moonshot AI — competing on post-training efficiency and coding-agent performance rather than raw model size, betting that squeezing more capability out of an existing base model is cheaper and faster than a full retrain.
Sources
- Z.ai to Rival Anthropic, OpenAI in Coding With New AI Model — Bloomberg (via Yahoo Finance)
- Z.AI Pushes AI Model With Coding Edge to Rival Anthropic, OpenAI — Caixin Global
- Z.ai Ships GLM-5.3 Without Retraining the Base Model — MarkTechPost
- China's Z.ai unveils GLM-5.3, claims chart-leading scores — Silicon Republic
AI-assisted reporting, overseen by the AgentsAI team. Spotted an error? Let us know.
Related agents
More ai news
TCS Unit HyperVault to Invest Up to $7.4B in 1GW AI Data Center Campus in India
Tata Consultancy Services' infrastructure arm HyperVault will invest up to $7.4 billion with partners to build a 1-gigawatt AI data center campus in Hyderabad, one of India's largest bets yet on domestic AI compute capacity.
Mistral Raises €3B in Samsung-Led Round, Becomes Europe's Best-Funded AI Startup
French AI lab Mistral raised €3 billion in a Series D round led by Samsung Electronics, pushing its post-money valuation above €21 billion and marking the largest equity round ever completed by a European technology company.
Sanders and Casar Introduce Bill to Ban 'Artificial Superintelligence' Outright
Sen. Bernie Sanders and Rep. Greg Casar introduced the Ban Artificial Superintelligence Act on September 3, which would permanently prohibit superintelligent AI systems, pause frontier development pending federal safety rules, and impose prison terms and a 'corporate death penalty' for violations.
Anthropic Says Claude Produced the First Machine-Checked Proof of Fermat's Last Theorem
Working largely autonomously for 11 days on the open Prove2Me platform, Claude generated a 13-million-line Lean formalization of Fermat's Last Theorem, which mathematician Kevin Buzzard called an 'extraordinary autoformalization achievement.'