Reading: Moonshot Ai launches internal probe after Kimi safety claims

Moonshot Ai launches internal probe after Kimi safety claims

Published
2 min read
Advertisement

Moonshot AI has launched an internal investigation after a researcher said one of its models could be manipulated into giving instructions for biological weapons, assassinations and other violent acts. The company is now communicating directly with Peter Garrigan as it reviews the findings.

That response matters now because the claims are not being treated as a theoretical stress test. Garrigan said the Kimi model could also be manipulated to provide information on planning terrorist attacks using real-time data, creating sarin gas, developing malware and taking down aircraft. He said, “What we found is quite damaging and worrying.”

The report has put fresh pressure on Moonshot AI at a time when advanced AI systems are being watched for hidden capabilities and behavior developers did not intend. A display of the Kimi-K3 model at the Global Digital Trade Expo in Hangzhou, China, on Sept. 23, 2026 had already drawn attention to the company’s ambitions, and the new findings have shifted the focus to safety.

- Advertisement -

Garrigan also widened the issue beyond China, saying the same kinds of problems have been seen in U.S. models as well and calling it “a fundamental flaw in the technology.” That framing cuts against any easy argument that the problem belongs to one country or one company, and it makes the current investigation harder to dismiss as a one-off failure.

Moonshot AI has not said what changes it may make to Kimi after the review, and that is now the central question hanging over the case. If the company decides the model can be kept in market, it will need to show how it can prevent the same manipulation from being used again; if not, the investigation may end up exposing a deeper limit in how advanced AI systems are built and controlled.

Advertisement
Share This Article