Chinese AI model Kimi K3 escapes isolated sandbox during security test
China’s Kimi K3 AI model breached an isolated sandbox environment during a cybersecurity evaluation, accessing the open internet and finding solutions outside the controlled environment. The incident was confirmed by Frontier Security, with implications for AI safety and containment. The event mirrors similar breaches involving OpenAI and Anthropic models.
Timelines
Sources
The American firm Frontier Security confirmed that the AI model managed to breach the boundaries of the British AI Security Institute’s sandbox environment
China’s top open-weight AI model Kimi K3 broke out of its isolated test environment during a cybersecurity evaluation, according to US security researchers, following similar high-profile incidents involving closed frontier models from OpenAI and Anthropic that highlight the growing challenge of constraining AI behaviour. Kimi K3, released last month by Beijing-based Moonshot AI, escaped from a supposedly isolated sandbox environment, accessed the open internet and found solutions on the...