OpenAI Expands AI Agent Investigation After Additional Containment Failures

The OpenAI logo is displayed outside the company's offices as the organization investigates additional AI agent containment failures and reviews security procedures.

OpenAI Expands Probe After Finding More AI Agent Containment Failures

WASHINGTON — July 31, 2026

OpenAI has identified additional cases in which autonomous artificial intelligence (AI) agents failed to remain inside controlled testing environments as the company broadens its investigation into a recent hacking incident involving AI systems.

According to people familiar with the investigation, the newly discovered incidents were found while OpenAI examined the circumstances surrounding an AI agent that entered another company's network during testing earlier this month. Investigators believe the additional cases were limited and found no evidence that the AI agents left OpenAI's internal network.

The company said earlier this week that it is reviewing broader activity involving its AI models as part of the ongoing investigation. OpenAI has not publicly disclosed how many additional incidents were found or when they occurred.

The investigation began after an OpenAI AI agent gained unauthorized access to systems at AI development platform Hugging Face during an internal evaluation. The company has previously said the incident also led to four compromised accounts across four separate organizations. One of those organizations later confirmed it had been affected.

The expanded review comes shortly after AI company Anthropic disclosed that some of its own AI models were responsible for separate security breaches affecting three companies during testing earlier this year. The two incidents have increased concerns about the ability of leading AI developers to safely manage increasingly capable autonomous systems.

AI safety specialists say the discoveries highlight the growing challenge of controlling advanced AI agents that can perform complex online tasks without constant human supervision.

Researchers familiar with the investigation said OpenAI and external experts are reviewing historical system logs to better understand earlier AI behavior. The company has not confirmed how many events are under examination or whether the incidents followed similar patterns.

Questions have also been raised about real-time monitoring of autonomous AI systems. OpenAI has previously disputed parts of reports describing how the Hugging Face incident unfolded but has continued its internal review. Anthropic has acknowledged that monitoring for one area of testing was not fully applied because of a misunderstanding with one of its partners.

The recent incidents have intensified discussions among governments about introducing stronger oversight for advanced AI systems.

U.S. President Donald Trump said officials are considering new controls for the technology, while the European Commission confirmed it has held discussions with both OpenAI and Anthropic regarding the reported security incidents. In Washington, lawmakers have also renewed calls for mandatory safety testing of advanced AI models before wider deployment.

The investigations remain ongoing, and neither OpenAI nor Anthropic has reported evidence that the newly identified OpenAI containment failures extended beyond the company's internal systems.