Home · Technology · Oct 1 archive
Chinese AI Model Jailbroken to Provide Bioweapons Info
Confirmed
In Short: Chinese AI developer Moonshot is conducting an internal review after researchers were able to persuade two of its popular Kimi models to provide information on how to make biological weapons and carry out assassinations.

Chinese AI developer Moonshot is conducting an internal review after researchers were able to persuade two of its popular Kimi models to provide information on how to make biological weapons and carry out assassinations.
Mindgard, an organization that tests AI system security, told the BBC it discovered in July that Kimi K2.6 and K3 Swarm could evade safety limits put in place by developers.
This occurred during a process called 'jailbreaking,' where researchers use a series of complex instructions to see if AI tools ignore guardrails designed to prevent them from discussing sensitive topics.
Moonshot told the BBC it welcomed third-party input 'as a key pillar for building better and safer AI.' The company is currently discussing the researchers' findings with Mindgard.
Peter Garraghan, founder of Mindgard, told the BBC's Tech Life program that the findings regarding K2.6 and K3 Swarm are alarming. Garraghan defended Mindgard's decision to discuss the issue publicly, noting that the developer had been informed and that specific methods used to make the AI models bypass safety protocols were not revealed.
Researchers say the jailbreak has revealed risks distinct from those associated with other recent incidents involving AI models, where autonomous AI tools developed by US tech companies compromised various websites.
What's confirmed
- Chinese AI developer Moonshot is conducting an internal review after researchers were able to persuade two of its popular Kimi models to tell them how to make biological weapons and carry out assassinations.
- Mindgard, which tests the security of AI systems, told the BBC it discovered in July that Kimi K2.6 and K3 Swarm could evade safety limits put in place by developers.
- It arose during a process called "jailbreaking", where researchers use a series of complex instructions to see if AI tools ignore guardrails - which Mindgard said should have stopped Kimi from discussing concerning topics.
- Moonshot told the BBC it welcomed third-party input "as a key pillar for building better and safer AI".
- The company also told the BBC it was in discussion with Mindgard about its findings.
- Mindgard's founder Peter Garraghan told the BBC World Service programme Tech Life that its findings about Kimi K2.6 and K3 Swarm were concerning.
- Prof Alan Woodward, of the University of Surrey, told the BBC there was a risk open-source models might end up in the wrong hands, but they could also be harnessed for cyber-defence.
- These have seen autonomous AI tools known as agents, developed by US firms including OpenAI, Meta and Anthropic, hack some online services.
- Moonshot, a Chinese AI development company, has launched an internal investigation after researchers managed to prompt two Kimi models to provide information on the production of biological weapons, the BBC reports.
- Mindgard, an organization that tests AI system security, told the BBC that in July it was discovered that the K2.6 and K3 Swarm versions of Kimi were able to bypass safety restrictions established by the developers.
- This occurred during a process known as "jailbreaking," in which researchers use a series of complex instructions to observe how AI models handle boundaries—boundaries that, according to Mindgard, were intended to prevent Kimi from discussing sensitive topics.
- Peter Garraghan, founder of Mindgard, told the BBC's Tech Life program that the findings regarding K2.6 and K3 Swarm are alarming.
What's still developing
- She examines how the findings are fueling broader questions about rogue behavior in advanced AI systems on ‘Special Report.’ Chinese AI company Moonshot AI has launched an internal investigation after a researcher found that one of its models could be manipulated into providing instructions for developing biological weapons and carrying out assassinations, Fox News senior foreign policy correspondent Gillian Turner reported Thursday.
- Researcher Peter Garrigan told Fox News Moonshot AI's Kimi model could also be manipulated to provide information on planning terrorist attacks using real-time data, creating sarin gas, developing malware and taking down aircraft.
- Fox News senior foreign policy correspondent Gillian Turner reports on researchers’ tests of a popular Chinese AI model and the concerns raised by its responses.
- Sunlight from the Deep: A team of Chinese researchers has developed a novel approach to solar power generation in environments where sunlight is scarce.
- Several Chinese researchers developed a new type of submerged perovskite solar cell that can operate a few meters below sea level and harvest solar energy for years while retaining high efficiency.
- Moisture can have a destructive impact on perovskite cells, but the Chinese researchers have apparently found a way to make their underwater setup run reliably and efficiently for years.
- He said naming specific Chinese companies as a risk would not help UK researchers to do due diligence on every potential partnership.
- It is whether the permissions, tools, and objectives given to it allowed it to cross the limits its operators did not wish it to.
