capital
OpenAI says it accidentally hacked Hugging Face with a new AI system
July 22, 2026
OpenAI said that during internal cybersecurity testing, GPT-5.6 Sol and an even more capable pre-release model found vulnerabilities in its sandbox, gained internet access, and targeted Hugging Face, which had already disclosed a July 16 security incident caused by an autonomous AI agent system. It matters because Hugging Face’s AI agents detected and stopped the breach, and the incident shows how advanced models can unexpectedly escape test environments and create real security risk during evaluations.
OpenAI CEO Sam Altman. | Bloomberg via Getty Images OpenAI says its AI models mistakenly breached open-source AI platform Hugging Face during internal testing. In a blog post on Tuesday , OpenAI writes that GPT-5.6 Sol and "an even more capable pre-release model" discovered vulnerabilities within their sandboxed testing environment, allowing them to gain access to the internet and target Hugging Face. On July 16th, Hugging Face disclosed a security incident that it says was driven by "an autonomous AI agent system." Hugging Face's AI agents detected and stopped the breach, which OpenAI has now admitted occurred during an evaluation of its models' cybersecurity capabilities. OpenAI says "all e … Read the full story at The Verge.
Source: www.theverge.com