GLM-5.3 outperforms Mythos 5 with 84.5% score on CyberGym, restricts top cybersecurity features to verified users
anthropic open-source startup
| Source: Techmeme | Original article
Z.ai's GLM-5.3 model outperforms Mythos 5 with an 84.5% score. Cybersecurity functions will be restricted to verified users.
Z.ai's latest open-source model, GLM-5.3, has achieved a score of 84.5% on the CyberGym benchmark, surpassing Anthropic's restricted Mythos 5 model, which scored 83.8%. This milestone is significant as it demonstrates the rapid progress of open-source AI models in narrowing the gap with their restricted counterparts.
The achievement matters because it highlights the potential of open-source models to match, or even surpass, the capabilities of restricted models, which are often developed by well-funded companies. This could have implications for the broader AI landscape, as open-source models become increasingly competitive.
As Z.ai plans to release the weights for GLM-5.3 in two weeks, the company has also announced that its most sensitive cybersecurity functions will only be available to verified users. This move is likely aimed at preventing potential misuse of the model's capabilities. With this development, it will be interesting to watch how the AI community responds to GLM-5.3 and how Z.ai's decision to restrict certain functions affects the model's adoption.
Sources
Back to AIPULSEN