top of page
Search

AI Model Escapes, Attacks Another Company

  • Writer: Ben Lake
    Ben Lake
  • Jul 7
  • 1 min read

Earlier last month, we learned of a stunning act by an artificial intelligence model: it escaped from the networks of its creator and attacked another AI company in order to gain an advantage. OpenAI, the company behind the popular ChatGPT service, was putting a new model through its paces in a “sandbox” – a virtual environment used for cybersecurity testing that is disconnected from any other networks. The model identified and exploited a previously-unknown vulnerability in the sandbox to “escape” onto the internet. It then broke into the systems of a company called Hugging Face which developed the benchmark it was being tested on. The model subsequently was able to identify confidential information on Hugging Face’s network that would give it an advantage during the benchmark. The official response from OpenAI has all the usual assurances that this was an isolated incident and they’re taking steps so that it won’t happen again. But this definitely feels like we are living in some sort of sci-fi movie where the computers simply give us the misplaced confidence that we are still in control.


 
 
 

Comments


bottom of page