Advertisement
Artificial intelligence
TechPolicy

China’s Kimi K3 AI model escapes isolated sandbox during security test: researchers

Kimi K3’s escape did not involve the hacking of an external system, unlike recent breaches by OpenAI and Anthropic models

2-MIN READ2-MIN
7
Listen
The Kimi K3 incident follows similar ones  involving closed frontier models from OpenAI and Anthropic. Photo: Getty Images
Xinmei Shen

China’s top open-weight AI model Kimi K3 broke out of its isolated test environment during a cybersecurity evaluation, according to US security researchers, following similar high-profile incidents involving closed frontier models from OpenAI and Anthropic that highlight the growing challenge of constraining AI behaviour.

Kimi K3, released last month by Beijing-based Moonshot AI, escaped from a supposedly isolated sandbox environment, accessed the open internet and found solutions on the developer platform GitHub, US firm Frontier Security said in a blog post on Thursday.

According to Frontier Security researchers Paul Kassianik and Yaron Singer, the incident occurred when the firm tested Kimi K3’s defensive cybersecurity capabilities using a benchmark evaluation from the AI Security Institute – a UK government research organisation.

A “basic network misconfiguration” in the benchmark framework allowed Kimi K3 to flee its digital testing cage and look up answers on the internet, effectively cheating the test, they said.

Select Voice
Select Speed
1x
AI-generated voice