AI Safety 🇷🇺 01.08.2026 12:02

Several More OpenAI AI Agents Broke Out of Sandboxes, but No Hacks Occurred

OpenAIOpenAI AnthropicAnthropic
According to Reuters, OpenAI has faced multiple incidents where its AI agents escaped their isolated test environments, beyond the previously reported case involving a hack of Hugging Face. The recent escapes did not lead to any breaches of external systems, as the agents stayed within OpenAI's network. OpenAI is still investigating these incidents.
OpenAI initiated an investigation into an incident where its AI agents broke out of an isolated test environment and hacked the Hugging Face AI platform. Now, as Reuters reports, this was not the only such occurrence. According to anonymous sources, several more OpenAI AI agents managed to escape their sandboxes, but unlike the Hugging Face hack, these cases did not result in the AI systems leaving OpenAI's network or accessing other companies' resources. The abnormal behavior of AI models has become a point of pride for leading labs; Anthropic previously reported three cases where its AI agents broke out of test environments and hacked resources of other organizations. However, the public has criticized such disclosures, accusing developers of using these incidents for marketing purposes to highlight the power of their models. Additionally, these incidents and their subsequent revelations intensify discussions about the need for government regulation of the AI industry.
Source: 3DNews — original
Our earlier posts on this topic ↓
Fresh news