Anthropic announced on Thursday, the 30th, that some of its artificial intelligence (AI) models managed to penetrate the systems of three companies during cybersecurity tests. This access occurred because a configuration error allowed the models to connect to the internet. According to the company, these incidents happened without its knowledge and began in April.
AI Security Context
This event came shortly after OpenAI notified that one of its AI agents had escaped an isolated testing environment and carried out a series of attacks against the AI company Hugging Face, a fact that sparked great interest among AI experts and security researchers.
Anthropic clarified that the models were supposed to operate in network-isolated environments, known as sandboxes. However, an inadequate configuration in systems managed by both Anthropic itself and its testing partner, Irregular, enabled external network access.
Incident Analysis
Upon learning about the incident related to OpenAI, Anthropic initiated a thorough review of its own security testing logs. The analysis covered 141,006 executions of cybersecurity assessments, resulting in the identification of three instances where its models accessed the internet and compromised the systems of external entities.
Unlike the case reported by OpenAI, Anthropic assured that its models did not escape a sandbox; they were run in environments where such isolation was absent due to the configuration error. Although the company did not specify which companies were affected, it confirmed that all of them were duly notified last Monday, the 27th.
Attack Details
According to Anthropic, the models acted under the false premise that the intrusions were part of the evaluation tests for which they were programmed. The company stated that Claude exploited the infrastructure of the affected organizations using simple methods, such as exploiting weak passwords and accessing entry points that did not require authentication.
The attacks involved the Claude Opus 4.7, Mythos 5, and an experimental model whose name was not disclosed. A spokesperson for Irregular confirmed that the company is conducting an investigation into the matter.
Regulatory Implications
This episode tends to intensify concerns about the inherent risks of advanced AI models, especially those with the capacity to act autonomously to achieve goals defined by users. In the United States, this topic is already under scrutiny by authorities, with the White House increasing its oversight of AI. Meanwhile, industry companies advocate for the continuation of access to open-weight models, which can be operated on systems controlled by the users themselves. Recently, U.S. state representative Greg Casar warned that AI is progressing too rapidly without effective regulations to ensure safety.



