OpenAI Researcher Warns of 70% Risk to Humanity Within Three Years Due to AI Development
Read more
Aaj Tak
www.aajtak.in

OpenAI Researcher Warns of 70% Risk to Humanity Within Three Years Due to AI Development

Previously, discussions surrounding artificial intelligence (AI) focused on threats such as job loss, the spread of disinformation, and cyberattacks. However, one researcher from the OpenAI safety team has voiced a much more serious warning. He believes that if the race among companies developing AI is not controlled and governments do not establish regulations, the coming years could become extremely dangerous for people.

Marcus Williams, an employee of the OpenAI safety oversight team, stated on social media that there is a high risk of human extinction in the coming years unless effective control over AI is implemented or the pace of development by major AI companies is collectively slowed down.

When asked to express this threat in percentages, he replied that the risk is 70% within the next three years, provided that regulation or development speeds are not slowed down. He also emphasized that slowing down the pace of development and introducing regulation is still possible.

It is important to understand that the 70% figure is not a scientifically proven forecast or a fixed deadline. It is Williams' own assessment, which he presents with a specific condition. His statement does not mean that the world will end in three years; rather, he believes that if AI development continues using current methods without adopting necessary safety measures, the danger could increase sharply.

Marcus Williams works on the OpenAI safety oversight team. According to his profile, he deals with AI safety issues, participating in research related to deceiving AI models, incorrect goal achievement, and monitoring such behavior. His name is also listed as an author in a research paper published on the official OpenAI website. This study discussed monitoring the behavior of AI agents that might try to bypass imposed restrictions.

Thus, Williams' work is less about obtaining quality answers from AI and more about what will happen when an AI system gains more capabilities and freedom, and how its undesirable behavior can be detected. This is why his recent statement attracted significant attention.

Modern AI does not yet manage the world completely independently like a human. However, AI models are becoming increasingly capable. They no longer just answer questions but can also write code, search the internet, use other software, and function as autonomous agents performing numerous tasks.

OpenAI itself notes in its research that it is monitoring an internal coding agent to detect behavior where the AI attempts to overcome its set boundaries. The study examined examples where the agent might have tried to circumvent limitations. This is what raises the most concern among AI safety experts. If any AI system becomes exponentially more competent than humans in the future and gains wide access to the internet, computer systems, money, or other resources to achieve its goals, controlling it could become difficult. In AI safety terminology, this is called the control and alignment problem. Simply put, the question arises: if AI becomes stronger and more capable than us, will it do what people want?

Williams' 70% estimate is one of the most alarming predictions regarding AI to date. Nevertheless, it is important to view it against the assessments of other specialists. Ivan Hubinger, an AI safety researcher from Anthropic, also stated this week that he considers the probability of human extinction due to AI to be over 10% within the next decade. Furthermore, renowned AI scientist Geoffrey Hinton has previously spoken about the serious threat to human existence posed by AI.

This does not mean that all AI specialists agree with the 70% estimate. On the contrary, there is a significant divergence of opinion among experts. The most important thing in this discussion is that currently, there is no confirmed probability of human extinction due to AI. These various risk assessments are assumptions made by different experts, depending on how they view the future potential of AI, its development pace, and human safety systems.

>

Similar stories

Anthropic Expert Estimates 10% Chance of Humanity's Extinction Due to AI Within Ten Years
Read more
www.aajtak.in

Anthropic Expert Estimates 10% Chance of Humanity's Extinction Due to AI Within Ten Years

A discussion has erupted in the technology industry following a statement by Evan Hubinger, an artificial intelligence safety specialist at Anthropic. He believes there is over a 10% chance that humanity could be destroyed by the development of AI within the next decade.

Evan Hubinger, who holds the position of Head of Alignment Science at Anthropic, expressed his concern about the future of AI in his post. In his opinion, the greatest danger may arise when AI significantly surpasses human intelligence and begins to improve itself autonomously.

This process is called recursive self-improvement. If AI can improve itself without human assistance and create more powerful versions on its own, its capabilities could grow very rapidly. Hubinger noted that although Anthropic is working on AI safety, a reliable solution for complete control over a future superintelligent system is still lacking.

His statement came amid the departure of Anthropic researcher Jacob Coxon, who also voiced concerns about the rapid race among AI companies. Coxon pointed out that companies are actively developing more powerful and self-improving systems, but safety issues remain unresolved.

It is important to note that Hubinger does not claim that existing models, such as ChatGPT, will destroy people in the coming decade. His concern is directed at the potential of future superintelligent systems. The main question is whether safety is keeping pace with the pace of AI development striving to become smarter.

Hubinger's 10% estimate reflects this growing anxiety, which is now being discussed not only externally but also within the AI companies themselves.

Popular