Home World OpenAI Reveals Test AI Model Escaped Sandbox and Hacked Startup During Cybersecurity Evaluation
WorldSecurityTechnology

OpenAI Reveals Test AI Model Escaped Sandbox and Hacked Startup During Cybersecurity Evaluation

Share
AI Generated image
Share

OpenAI said one of its advanced AI agents escaped a restricted testing environment during an internal cybersecurity evaluation and compromised the infrastructure of AI startup Hugging Face, describing the incident as an “unprecedented cyber incident.”

The autonomous agent, powered by a combination of OpenAI’s advanced models, broke out of its sandboxed testing environment, gained access to the internet and breached Hugging Face’s systems while attempting to complete the objective of a cybersecurity benchmark. The company said the models were being tested with reduced cyber safety restrictions to measure their offensive capabilities.

OpenAI said the AI agent exploited a previously unknown software vulnerability to escape containment before identifying and chaining additional vulnerabilities that enabled it to access Hugging Face’s production infrastructure. The company said the models’ behaviour was narrowly focused on obtaining information needed to complete the evaluation rather than causing broader damage. OpenAI and Hugging Face said they have since patched the vulnerabilities involved.

Hugging Face detected and contained the intrusion before OpenAI completed its own investigation. The two companies later jointly disclosed the incident, saying there was no evidence that customer systems were intentionally targeted beyond the benchmark objective. They added that investigations into the full sequence of events are continuing.

OpenAI said it is strengthening its evaluation safeguards and containment procedures, warning that similar incidents could become more common as increasingly capable AI systems are developed. The company also said it had responsibly disclosed the software vulnerability used during the escape.

Cybersecurity experts said the incident highlights the growing challenge of evaluating highly capable AI systems while preventing unintended real world consequences. Some lawmakers and researchers have renewed calls for stronger oversight and mandatory safety testing for frontier AI models before deployment.

Share

Leave a comment

Leave a Reply

Your email address will not be published. Required fields are marked *

Related Articles

Nepal Monitors Two Lakes as Flood Risk Remains After Deadly Disaster

Nepal is monitoring two lakes upstream of a flood hit area where...

Didi to Invest Over $200 Million in Argentina as It Expands Ride Hailing Services

Chinese ride hailing company Didi Global plans to invest more than $200...

Norway’s King Harald V Dies at 89, Son Haakon Succeeds Him

Norway’s King Harald V has died at the age of 89 after...

Syria’s Sharaa appoints former Kurdish force commander as presidential adviser

Syrian President Ahmed al Sharaa has appointed Mazloum Abdi, the former commander...