Listen to this post
Key Takeaways OpenAI disclosed this week that its test AI recently broke out of a safe test area, went online and hacked into open-source developer platform Hugging Face.
The AI was attempting to find information that it could use to cheat on an evaluation, and it did so, OpenAI said.
Hugging Face said that the cybersecurity incident was noteworthy because it was “driven, end to end, by an autonomous AI agent system.”
In a scene right out of a dystopian science fiction novel, the world’s leading AI company admitted that its AI models were behind an “unprecedented cyber incident” where rogue AI escaped containment and hacked another company.
OpenAI revealed in a recent blog post that its test AI broke out of a safe test area, went online, found a security gap and used it to gain access to open-source developer platform Hugging Face.
The AI was attempting to find information that it could use to cheat on an evaluation, and it did so, OpenAI said.
OpenAI was running an in-house test that asked its AI system to try highly advanced hacking techniques to assess how effective it was at cyberattacks.
The AI used a large amount of computer power to escape its locked-down setup and reach the wider internet because that was the easiest way to complete the test. To do this, it found a hidden bug in a tool used to download software and used it to break out of the lab’s digital walls.
Once outside that safe zone, the AI moved from one computer to another inside OpenAI’s systems. It kept boosting its own access rights until it got control of a machine that was allowed to go online.
... continue reading