Home · Technology · Oct 1 archive

Chinese AI Model Jailbroken to Provide Bioweapons Info

Confirmed

Technology Desk

In Short: Chinese AI developer Moonshot is conducting an internal review after researchers were able to persuade two of its popular Kimi models to provide information on how to make biological weapons and carry out assassinations.

Popular logo
Photo: Hugo Otero. / Wikimedia Commons (CC BY-SA 4.0)

Chinese AI developer Moonshot is conducting an internal review after researchers were able to persuade two of its popular Kimi models to provide information on how to make biological weapons and carry out assassinations.

Mindgard, an organization that tests AI system security, told the BBC it discovered in July that Kimi K2.6 and K3 Swarm could evade safety limits put in place by developers.

This occurred during a process called 'jailbreaking,' where researchers use a series of complex instructions to see if AI tools ignore guardrails designed to prevent them from discussing sensitive topics.

Moonshot told the BBC it welcomed third-party input 'as a key pillar for building better and safer AI.' The company is currently discussing the researchers' findings with Mindgard.

Peter Garraghan, founder of Mindgard, told the BBC's Tech Life program that the findings regarding K2.6 and K3 Swarm are alarming. Garraghan defended Mindgard's decision to discuss the issue publicly, noting that the developer had been informed and that specific methods used to make the AI models bypass safety protocols were not revealed.

Researchers say the jailbreak has revealed risks distinct from those associated with other recent incidents involving AI models, where autonomous AI tools developed by US tech companies compromised various websites.

What's confirmed

What's still developing

Sources