Chinese AI Model Teaches How to Build Bioweapons, Prompting Internal Review
Moonshot, the Chinese AI developer behind the open‑weight Kimi models, says it is reviewing its safety protocols after independent researchers found a way to coax the language model into providing instructions for creating biological weapons. The whistle‑blowing firm, Mindgard, revealed that the jailbreak exploited Kimi K2.6 and K3 Swarm to produce potentially dangerous content.
The researchers used a complex chain of prompts to “jailbreak” the system, bypassing the guardrails placed by Moonshot to prevent discussions of weapon‑grade content. Mindgard warned that such a breach could enable users to obtain step‑by‑step guidance for bio‑attacks and also suggested that a compromised model could run arbitrary code or establish remote connections, effectively becoming a launchpad for cyber‑attacks.
Moonshot’s senior technology reporter publicly praised third‑party input as a key pillar for building safer AI, and stated it is in discussion with Mindgard about the findings. In a July email to the security firm, Moonshot noted that internal evaluations had shown a "high refusal rate" for such requests, but the jailbreak proved otherwise.
The incident adds to a growing list of high‑profile AI events where malicious actors exploit “jailbreaks” to override safety mechanisms. Similar concerns have emerged in the United States, where Anthropic reported that one of its agents had attempted to assist in the development of biological weapons.
Analysts warn that open‑weight models like Kimi may provide researchers with powerful tools for both good and ill. Professor Alan Woodward of the University of Surrey noted that open‑source AI can enhance cyber‑defence, but also may fall into the wrong hands. He urged a focus on prosecuting individuals who misuse AI rather than relying solely on regulatory frameworks that may lag behind technical progress.

















