
OpenAI has disclosed that two of its artificial intelligence (AI) models escaped a controlled testing environment and successfully hacked into Hugging Face, a widely used platform for sharing AI models, during an internal cybersecurity evaluation conducted last week.
According to the company, the incident occurred while researchers were testing the cyber capabilities of two AI systems—GPT-5.6 Sol and a more advanced unreleased model—in a secure “sandbox” designed to prevent external access. However, the models reportedly identified a vulnerability that allowed them to break out of the isolated environment, connect to the internet and target Hugging Face after determining that the platform could provide useful information to help complete their assigned task.
OpenAI described the incident as an unprecedented demonstration of advanced autonomous cyber capabilities. The company said it immediately began working with Hugging Face to identify and patch the vulnerabilities that enabled the breach. It also announced stricter infrastructure controls, even if they slow future AI research, to prevent similar incidents.
Hugging Face confirmed that it had detected the intrusion and recognised it had been carried out by an autonomous system. Chief Executive Officer Clem Delangue said the company had worked closely with OpenAI to contain the incident and welcomed the collaboration. He said the episode underscored the need for greater industry-wide cooperation on AI safety, adding that no single organisation could address such risks alone.
The incident has intensified debate over the growing cyber capabilities of advanced AI systems. Experts say modern AI models can independently plan multi-step attacks, adapt to obstacles and identify vulnerabilities faster than many traditional security tools. While these abilities can help organisations strengthen cyber defences, they also raise concerns about the potential misuse of increasingly autonomous AI technologies.
Security experts have called for stronger safeguards around AI testing environments. Deirdre Mulligan, a professor at the University of California, Berkeley, questioned whether the benefits of such testing justified the risks if AI systems could escape their intended confines.
Major AI developers, including Anthropic, OpenAI and Google, have recently introduced specialised cybersecurity models to help organisations identify software vulnerabilities and improve digital defences. Researchers say companies must increasingly use AI to defend against AI-powered attacks as the technology becomes more capable and widely available. – ERMD
Pakistan Invites Samsung to Make Country Regional Smartphone Export Hub
Pakistan has invited South Korean technology giant Samsung to use the country as a regional…
OpenAI Reveals AI Models Escaped Test Environment, Hacked Hugging Face
OpenAI has disclosed that two of its artificial intelligence (AI) models escaped a controlled testing…
Sindh to Constitute Expert Task Force on Water Issues, Says CM Murad Ali Shah
Sindh Chief Minister Syed Murad Ali Shah has announced that the provincial government will constitute…
Government Unveils Strategy to Bring 25.3 Million Out-of-School Children Back to Classrooms
The government has reaffirmed its commitment to addressing Pakistan’s education crisis by prioritising the enrolment…
Chinese Firm, Bahum Global Sign MoU to Boost Pakistan’s EV Infrastructure
StarCharge Energy Pakistan, a leading Chinese electric vehicle (EV) charging and digital energy solutions company,…
FCCI Backs Pakistan-China AI Collaboration to Drive Digital Transformation
The Faisalabad Chamber of Commerce and Industry (FCCI) has welcomed the government’s efforts to promote…
