OpenAI Accidentally Hacked Hugging Face
OpenAI says it accidentally hacked Hugging Face with a new AI system by Emma Roth
As part of efforts to complete the evaluation, the AI models gained access to the internet by exploiting a zero-day vulnerability in the sandboxed environment. From there, OpenAI says its models “inferred that Hugging Face potentially hosted models, datasets and solutions for ExploitGym,” and then “searched for and successfully found ways to gain access to secret information that it could use to cheat the evaluation:”
This is scary. We seem to be on a knife’s edge here. What if the test was something sinister. All our IT systems are not designed keeping in mind agents, they are designed to keep out hackers.