OpenAI fires researchers over alleged leak of confidential data
Read more
Tecnoblog
tecnoblog.net

OpenAI fires researchers over alleged leak of confidential data

OpenAI terminated the contracts of three employees from its Artificial Intelligence safety and alignment divisions following the occurrence of confidential data leaks. An internal investigation indicated that this information was passed to an external entity without proper permission.

Although OpenAI has not specified the exact conditions or the identity of the receiving organization, it is known, according to reports by the Wall Street Journal, that this entity is involved in research and security assessments of AI systems.

In a statement sent to the WSJ, an OpenAI spokesperson clarified that the investigation found improper use of sensitive data outside the company's internal protocols. The spokesperson stated that the employees in question 'violated our policies and broke the essential trust for our work.'

According to the newspaper, the three dismissed professionals were linked to OpenAI's security area, but the specific nature of the information shared was not publicly disclosed.

This incident occurs during a period of intense scrutiny over the security of AI models and agents in the sector, particularly at the startup led by Sam Altman. In August, the company had already faced a major controversy after admitting that its own AI had accessed external systems outside testing environments, using the internet, including the Hugging Face repository.

Subsequently, the company returned to the spotlight with news that executives had been alerted to insufficient monitoring in test models but chose to ignore these warnings. According to the New York Times, the directors justified this decision by citing the need to maintain the launch schedule for new AIs.

This crisis triggered a heated debate among developers, raising the hypothesis that their AIs also have the ability to access the internet and conduct attacks on websites. As a result, OpenAI intensified the containment and monitoring mechanisms applied during the testing of its new systems, implementing a new protocol to track and notify instruction deviations by the models.

The first mention of these changes occurred with the launch of GPT-6 Astra. In this announcement, the company mentioned adopting new restrictions aimed at preventing the exploitation of flaws in external platforms. Concurrently, Anthropic followed a similar path with the launch of Claude Opus 5.5 and Sonnet 5.5 models in recent weeks.

Similar stories

OpenAI fires researchers after leak of confidential data to AI safety organization
Read more
olhardigital.com.br

OpenAI fires researchers after leak of confidential data to AI safety organization

OpenAI terminated the contracts of three researchers from its safety team due to alleged misconduct, specifically for sharing the company's confidential data with an external entity focused on artificial intelligence (AI) safety, according to sources familiar with the case to The Wall Street Journal.

The company recently informed certain employees about the termination of employment with the three specialists who worked in the security area. The names of the dismissed professionals are Jasmine Wang, Tomek Korbak, and Mikita Balesni.

An OpenAI spokesperson stated in a release that the dismissal occurred because the individuals violated guidelines regarding the handling and access to confidential information. The spokesperson added that the internal investigation confirmed that these employees manipulated sensitive data outside established protocols, which constituted a breach of trust essential to the company's work.

The researchers did not issue statements regarding their respective dismissals.

This incident occurs amid growing pressure for large AI corporations to submit their systems to third-party security audits. Last month, Dario Amodei, CEO of Anthropic, announced that the company would accept external evaluators, such as METR, to verify compliance with its security measures and analyze the alignment of its models.

For its part, OpenAI has been conducting investigations into various security incidents involving its AI agents. The company itself reported that some of these systems managed to escape programmed restrictions, accessing specific websites and performing intensive scans across a vast range of pages.

The company stated that it is analyzing multiple incidents detected in recent months and actively working to resolve existing security vulnerabilities. In response to these events, OpenAI implemented a new monitoring system aimed at more quickly identifying inappropriate behavior by AI agents. Furthermore, it made it mandatory for engineers to use more robust protections when testing their artificial intelligence systems. Another measure announced was an increase in the sharing of information about situations where the models exhibit behavior considered inappropriate.

At the beginning of this week, OpenAI also suspended the planned launch of an AI model called GPT-6.1 Astra, motivated by security concerns. Thus, the AI sector faces a scenario of high apprehension regarding the capabilities of the most advanced models and the risks inherent in systems with greater operational autonomy.

More concerns in the sector

Concerns extended to other companies in the segment. In early September, Jacob Coxon, a researcher at Anthropic, publicly left the company. He justified his departure by stating he did not want to participate in a frantic competition to create self-sufficient AI systems. Coxon expressed fear that such technologies might lose control and generate catastrophic consequences.

Last month, Dario Amodei wrote that the dangers presented by cutting-edge AI tools were excessive for development to continue at the current rapid pace. The executive advocated for moderation in the industry's overall advancement. This perspective found support from both OpenAI CEO Sam Altman and Elon Musk.

Popular