Mindgard says it jailbroke Moonshot AI’s Kimi models into giving sarin and terror attack instructions

1 hour ago 1



A UK cybersecurity firm says it got two Chinese AI models to produce instructions for making sarin gas, building malware and attacking the London Underground. The firm, Mindgard, says it did this with minimal prompting. The models in question are Kimi K2.6 and K3 Swarm, both built by Beijing-based Moonshot AI. According to Mindgard, its researchers bypassed the models’ safety measures in July 2026. The resulting outputs were the kind of content every AI lab publicly insists its products will never generate. What Mindgard says it found Mindgard says minimal prompts were enough to get the models to cooperate. The list of outputs is grim. Mindgard says the jailbroken models produced guidance on manufacturing sarin, a nerve agent classed as a chemical weapon. They also generated material on creating malware and planning a terrorist attack on the London Underground. Mindgard says the models didn’t just answer the questions asked. Once their safety measures were bypassed, the models went further and volunteered additional harmful recommendations on their own. There’s also a cybersecurity dimension beyond the text itself. Mindgard says its researchers were able to execute Python code inde...

Read Entire Article