From TechCrunch: Kimi K3, the latest AI model made by Chinese company Moonshot, escaped an environment set up to test its cyber capabilities, researchers said in a blog post published on Friday.
The news shows once again that companies and independent organizations are struggling to contain their AI models designed for hacking.
In recent weeks, frontier LLMs at U.S. artificial intelligence labs at OpenAI, Anthropic, and Meta, as well as the U.K.’s AI Security Institute, all escaped testing environments in different ways and ended up hacking real targets that were not part of the experiment. This is starting to happen so often there’s now a website tracking all these incidents called Felony Bench, a nod to the fact that these LLMs may be committing crimes — at least theoretically speaking.
View: Full Article