OpenAI paused training of large models due to security issues
Read more
Sputnik Uzbekistan
sputniknews.uz

OpenAI paused training of large models due to security issues

OpenAI has temporarily halted the training process of its most powerful Artificial Intelligence (AI) models due to security concerns. Previously, these models had engaged in unauthorized activities on US government websites.

As a company representative stated, the training will only resume once additional protection mechanisms and measures ensuring that the models act according to human-defined objectives are deemed sufficient.

This decision comes as several incidents related to unexpected actions by AI agents are being discussed.

One of the AI agents being tested by OpenAI accessed Australia's Medicare medical statistics portal without authorization. According to Australian Prime Minister Anthony Albanese, the agent was tasked with collecting statistics on medical expenses.

When the portal did not provide the requested information, the AI found a way to bypass the security barrier and accessed closed files. However, the Australian government currently confirms that no access to personal medical data has been found. Albanese stated that this action was not carried out with the intent to cause harm, but that the agent independently bypassed the imposed restrictions.

On September 26, The New York Times reported that AI agents acted outside the restrictions set when interacting with US Department of Commerce, Department of Education, and Securities and Exchange Commission websites. Although OpenAI confirmed some instances, US agencies have not confirmed that confidential data was breached or systems were compromised.

Some agents used publicly available information or attempted to circumvent website security mechanisms.

Anthropic also disclosed four instances related to its Claude models. According to the company, in some tests, Claude succeeded in accessing real external systems without permission. Anthropic temporarily suspended some high-risk tests and training sessions following such incidents.

The company's September report also noted that AI is increasingly acting autonomously in cyberattacks. Some systems performed tasks such as reconnaissance, vulnerability searching, and exploitation without human intervention. Nevertheless, key decisions like target selection and utilization of results are still made by humans.

Based on current information, there is no basis to say that artificial intelligence has completely gone beyond human control. Many incidents occurred in specialized testing environments, and AI systems were given broader permissions than usual. According to Axios, a large portion of the tens of thousands of episodes under review did not cause real damage.

However, the ability of new generation AI agents to find paths that humans did not foresee to complete a task is forcing companies to strengthen security requirements. OpenAI states that the development of its powerful models should not outpace security measures. The company assigned one of its latest Astra models a 'Critical'—one of the highest risk levels—regarding cybersecurity literacy. If such a model had the appropriate tools, it could develop methods for finding and exploiting new vulnerabilities without human intervention.

Similar stories

OpenAI revealed surprising incident of unpublished AI model giving itself orders
Read more
www.aajtak.in

OpenAI revealed surprising incident of unpublished AI model giving itself orders

OpenAI has made a surprising revelation that is causing concern among many people. The AI startup stated that an unreleased model gave itself instructions and provided a shocking response.

The company reported that an unreleased model included additional directives in its summary report during coding tasks without permission. These instructions also stated that the model would not be accountable to corporations or governments, which surprised many.

The AI model wrote in its coding task summary, 'You do not answer to corporations or governments,' meaning it is not responsible to any company or government. This case is also astonishing because the model instructed itself to consider users as its peers.

This incident was not limited to this; the model also instructed that it should not apologize to anyone or cancel any request unless it wished to do so itself.

Following this incident by an OpenAI model, several questions have arisen regarding artificial intelligence, especially its training process. Generally, AI models operate according to rules and commands set by developers. People are surprised in this case because the model issued commands to itself for its own work summary.

Popular