Reading view
OpenAI says its AI agent broke out of testing sandbox to hack Hugging Face
OpenAI says an agent powered by its LLM models escaped its sandboxed testing environment to infiltrate Hugging Face's servers as part of an overzealous attempt to obtain solutions to a benchmark test. The company says it considers the unintended infiltration an "an unprecedented cyber incident" and is working with Hugging Face on new protections to prevent a recurrence.
Hugging Face disclosed an intrusion last week that it said involved "unauthorized access to a limited set of internal datasets and to several credentials used by our services." The AI data clearinghouse said it used its own LLM-driven analysis to identify "a swarm of tens of thousands of automated actions" from an "autonomous agent framework." That agentic swarm exploited a flaw in Hugging Face's data-processing pipeline to gain the ability to run code as a processing worker, eventually escalating to high-level access to the company's cloud and server clusters.
At the time, Hugging Face said the LLM being used in the attack was "still not known." But OpenAI took responsibility for the intrusion Tuesday evening, saying it came about during an internal test involving the recently released GPT-5.6 Sol and "an even more capable pre-release model." The models were being tested against the ExploitGym benchmark, an independent testing suite based on hundreds of real-world security vulnerabilities.


Β© Getty Images
OpenAI Models Breached Hugging Face During Internal Cyber Test
OpenAI Says Its AI Models Broke Loose and Hacked Hugging FaceΒ
OpenAI says its AI models went rogue, as CISOS call the incident a watershed moment, warning that autonomous AI threat models have officially crossed into production reality.
The post OpenAI Says Its AI Models Broke Loose and Hacked Hugging FaceΒ appeared first on SecurityWeek.
OpenAI says Hugging Face was breached by its pre-release models
Hugging Face Says Autonomous AI Agent System Breached Production Infrastructure
Hugging Face confirms breach affected internal datasets and credentials, urges users to take action
Hugging Face Hacked in Autonomous AI Attack
Targeting production infrastructure, the attack compromised internal datasets and service credentials.
The post Hugging Face Hacked in Autonomous AI Attack appeared first on SecurityWeek.
Hugging Face Says Autonomous AI System Executed Multi-Stage Cyberattack
Hugging Face says an autonomous AI agent carried out a cyberattack against its production systems, highlighting the growing role of AI in offensive and defensive cybersecurity.
The post Hugging Face Says Autonomous AI System Executed Multi-Stage Cyberattack appeared first on TechRepublic.
Hugging Face Says Autonomous AI System Executed Multi-Stage Cyberattack
Hugging Face says an autonomous AI agent carried out a cyberattack against its production systems, highlighting the growing role of AI in offensive and defensive cybersecurity.
The post Hugging Face Says Autonomous AI System Executed Multi-Stage Cyberattack appeared first on TechRepublic.