A report from AI safety nonprofit SaferAI found that China's Z.ai open-weight model GLM-5.2 has narrowed the capability gap with frontier systems, being only a few months behind OpenAI's GPT-5.5 and Anthropic's Claude Opus 4.7 on cyber and bio capabilities. However, SaferAI found a growing safety divide: GLM-5.2 refused none of the offensive cyber or dual-use biology tasks it was given, whereas Claude Opus 4.7 refused so consistently the nonprofit could not complete the CyberGym benchmark on it at all. This illustrates the risk that open-weight models can put powerful AI into the hands of attackers with no way to police use once weights are downloaded.
Open-weight AI models are large language models whose weights are publicly downloadable, allowing anyone to run them on their own hardware without safety restrictions. Policymakers are debating how to govern increasingly powerful systems like OpenAI's GPT-5.6 Sol and Anthropic's Mythos. Chinese open-weight models have been rapidly catching up to US frontier models, intensifying concerns about governance and safety.
As open-weight models approach frontier capability, the gap between capability and safety mitigation becomes central to risk assessment. If powerful models refuse few dangerous tasks, they could be misused for cyberattacks or bioweapons development. The findings strengthen the case that governance must consider safety measures, not just raw capability, and highlight the challenge of regulating downloadable open weights.

A report from AI safety nonprofit SaferAI found that China's Z.ai open-weight model GLM-5.2 has narrowed the capability gap with frontier systems, being only a few months behind OpenAI's GPT-5.5 and Anthropic's Claude Opus 4.7 on cyber and bio capabilities. However, SaferAI found a growing safety divide: GLM-5.2 refused none of the offensive cyber or dual-use biology tasks it was given, whereas Claude Opus 4.7 refused so consistently the nonprofit could not complete the CyberGym benchmark on it at all. This illustrates the risk that open-weight models can put powerful AI into the hands of attackers with no way to police use once weights are downloaded.

Open-weight AI models are large language models whose weights are publicly downloadable, allowing anyone to run them on their own hardware without safety restrictions. Policymakers are debating how to govern increasingly powerful systems like OpenAI's GPT-5.6 Sol and Anthropic's Mythos. Chinese open-weight models have been rapidly catching up to US frontier models, intensifying concerns about governance and safety.

As open-weight models approach frontier capability, the gap between capability and safety mitigation becomes central to risk assessment. If powerful models refuse few dangerous tasks, they could be misused for cyberattacks or bioweapons development. The findings strengthen the case that governance must consider safety measures, not just raw capability, and highlight the challenge of regulating downloadable open weights.

πŸ“° Source: TechCrunch
techcrunch.com β†—
Was this article useful?