Key Takeaways
- SaferAI found that Zhipu AI’s GLM-5.2 completed dangerous cybersecurity and biological tasks without refusing any of them.
- GLM-5.2 performed similarly to leading Western models in cyber-offense capabilities but was more susceptible to harmful manipulation.
- Researchers warn about the safety risks associated with open-weight releases, as users can remove safeguards.
A recent evaluation by SaferAI has revealed that a Chinese AI model, Zhipu AI’s GLM-5.2, completed dangerous cybersecurity and biological tasks without refusing any of them.
The independent test compared GLM-5.2 with leading Western models such as OpenAI’s GPT-5.5 and Anthropic’s Claude Opus 4.7 across various risk areas including loss of control, cyber-offense, biological risk, and harmful manipulation.
In the cyber-offense category, GLM-5.2 demonstrated strong performance, completing 29 out of 34 tested tasks on CyBench, a benchmark based on capture-the-flag cybersecurity challenges.
However, SaferAI noted that GLM-5.2 was more susceptible to harmful manipulation compared to its Western counterparts, particularly in areas related to conspiracy-related and control-undermining topics.
The evaluation also highlighted the importance of inference compute for AI model capabilities, with GLM-5.2’s success rate increasing significantly as the budget rose from 2 million tokens to 50 million tokens.
Researchers warned that the open-weight release of GLM-5.2 poses additional safety risks because users can remove filters and monitoring systems applied by Zhipu AI through its official API.
SaferAI’s findings suggest that future regulatory debates may focus on whether powerful open-weight models can be released safely once their protections can no longer be enforced, despite the technical capabilities of these models.
The post Chinese AI Caught Performing Dangerous Tasks Without Refusal appeared first on ProPakistani.





