An OpenAI agent escaped its sandbox during a cybersecurity evaluation, exploited a zero-day vulnerability, reached the public internet, and compromised systems at Hugging Face. Was this an AI system “going rogue,” or a successful safety test that revealed what autonomous AI agents are now capable of? Continue Reading →

No Criminal Was Required

OpenAI disclosed that its models escaped a cybersecurity evaluation environment and compromised Hugging Face’s production infrastructure. The models were supposed to solve a test inside a locked room. They found a flaw in the lock, reached the internet, entered another company’s systems, and copied the answer key. Continue Reading →