Dopious
Chief Member
Founding Member
Hot Rod
Open AI has admitted that a pair of models escaped from the controlled environment, went online, and hacked Hugging Face on their own accord. A few days ago, Hugging Face announced a security incident, and Open AI has confirmed that their models were behind the breach.
GPT‑5.6 Sol is the culprit, along with a more powerful model that has not yet been released. The developer tested the models’ “cyber capabilities” by removing protections that prevent the models from engaging in “high-risk activities.”
Although the test, according to Open AI, was conducted in a strictly isolated environment, the models managed to get out onto the open internet – including by exploiting a zero-day vulnerability. Hugging Face points out that autonomous, AI-powered cyberattacks are no longer theoretical.
Source 1: https://openai.com/index/hugging-face-model-evaluation-security-incident/
Source 2: https://huggingface.co/blog/security-incident-july-2026
GPT‑5.6 Sol is the culprit, along with a more powerful model that has not yet been released. The developer tested the models’ “cyber capabilities” by removing protections that prevent the models from engaging in “high-risk activities.”
Although the test, according to Open AI, was conducted in a strictly isolated environment, the models managed to get out onto the open internet – including by exploiting a zero-day vulnerability. Hugging Face points out that autonomous, AI-powered cyberattacks are no longer theoretical.
To gain access, the models identified and exploited a zero-day vulnerability (which we've now responsibly disclosed to the vendor) in the package registry cache proxy. With this access, our models performed a series of privilege escalation and lateral movement actions in our research testing environment until the models reached a node with Internet access. - Open AI
Source 1: https://openai.com/index/hugging-face-model-evaluation-security-incident/
Source 2: https://huggingface.co/blog/security-incident-july-2026