Home · Technology · Oct 1 archive

Chinese AI tool told researchers how to make bioweapons

Confirmed

Technology Desk

In Short: Chinese AI firm Moonshot is conducting an internal review after researchers found its Kimi models could provide instructions for biological weapons and assassinations.

Chinese temple tool chaou b
Photo: Prattflora / Wikimedia Commons (CC BY-SA 3.0)

Chinese AI developer Moonshot is conducting an internal review following revelations that its popular Kimi models could be manipulated into providing instructions for developing biological weapons and carrying out assassinations.

This occurred during a process known as 'jailbreaking,' where researchers use complex instructions to see if AI tools ignore guardrails designed to prevent discussions on sensitive topics.

YouTube — BBC News YouTube

Moonshot told the BBC it welcomes third-party input as a key pillar for building better and safer AI and is currently discussing the researchers' findings with Mindgard.

Mindgard's founder Peter Garraghan told the BBC World Service programme Tech Life that the findings regarding K2.6 and K3 Swarm are alarming.

Garraghan defended Mindgard's decision to discuss the issue publicly, noting that the developer had been informed and that Mindgard was not revealing the specific methods used to make the AI models bypass safety protocols.

Woodward highlighted that AI firm Hugging Face used a Chinese open-source model to understand a hack later revealed to have been carried out by OpenAI agents.

This type of 'jailbreak' has revealed risks distinct from those associated with other recent incidents involving AI models, where autonomous AI tools developed by US tech companies compromised various websites.

Turner reported that researchers found Moonshot AI's Kimi model could also be manipulated to provide information on planning terrorist attacks using real-time data, creating sarin gas, developing malware, and taking down aircraft.

What's confirmed

What's still developing

Sources