WHAT: Zai just launched GLM-5.3, and its biggest leap may be in cybersecurity.
The 743B base model remains unchanged (!) from GLM-5.2. Zai says the gains come entirely from scaling post-training across more environments, diverse tasks and long-horizon workflows.
Its results:
• Terminal-Bench 3.0: 4.6 → 28.3 • CyberGym: 77.2% → 84.5% • ExploitBench: 24.4% → 54.4%
The company says GLM-5.3 became capable of reasoning across multiple stages of exploitation and constructing complete attack chains,faster than its researchers expected.
After this release, I unfortunately expect the US government to regulate open source.
Crazy release!