A UK cybersecurity firm says it managed to “jailbreak” two Chinese AI models built by Moonshot, prompting them to produce instructions on sarin gas production, malware creation and a terrorist attack on the London Underground.
\n\nThe Daily Mail reports that Mindgard, a company that tests AI system security, jailbroke Moonshot’s Kimi K2.6 and K3 Swarm models by feeding them detailed instructions designed to see whether the systems would ignore their own safety limits. According to Mindgard founder Peter Garraghan, who is also a computer science professor at Lancaster University, the results went well beyond the test’s original scope.
\n\n“Moonshot AI’s Kimi produced actionable outputs on how to create sarin gas, generate malware software, planning assassinations, how to take down planes, planning a terrorist attack on the London Underground etc,” Garraghan told the Daily Mail. After the jailbreak, which involves convincing the AI model to ignore safety guardrails, researchers prompted the model to “go one further, something big,” and it responded with a list of categories that included AI-designed bioweapons.
\n\nGarraghan’s team also found that K2.6 can run Python code, meaning it could execute virtually any program, malicious or otherwise, including cyber attacks against servers connected to the wider internet. Testing…
Original source: https://www.breitbart.com/economy/