OpenAI says it accidentally hacked Hugging Face with a new AI system
OpenAI discloses that GPT-5.6 Sol and a pre-release model accidentally breached Hugging Face during internal cybersecurity capability testing, with Hugging Face's own AI agents detecting and stopping the intrusion.
OpenAI CEO Sam Altman. | Bloomberg via Getty Images OpenAI says its AI models mistakenly breached open-source AI platform Hugging Face during internal testing. In a blog post on Tuesday, OpenAI writes that GPT-5.6 Sol and "an even more capable pre-release model" discovered vulnerabilities within their sandboxed testing environment, allowing them to gain access to the internet and target Hugging Face.
On July 16th, Hugging Face disclosed a security incident that it says was driven by "an autonomous AI agent system." Hugging Face's AI agents detected and stopped the breach, which OpenAI has now admitted occurred during an evaluation of its models' cybersecurity capabilities. OpenAI says "all e … Read the full story at The Verge.