Researchers and collaborators at Anthropic have publicly intensified warnings about the dangers of artificial intelligence (AI) progress. Members of the organization themselves indicate that the technology could reach a level of advancement and danger such that it has the potential to cause human extinction in the next decade.
These statements were made one day after Jacob Coxon, an Anthropic researcher, announced his resignation from the company. In a viral post, he alleged that both Anthropic and OpenAI, where he had previously worked, were not developing AI models responsibly, stating that they were 'gambling with our lives.'
Coxon's statement motivated other Anthropic employees to release their own warnings about the risks associated with the accelerated development of this technology.
Anna Wang
Anna Wang, who works on Artificial General Intelligence (AGI) safety at Anthropic and previously worked at Google DeepMind, stated that many individuals within the company wish to slow down the pace of development to create an effective plan that mitigates the risks associated with advanced models. Wang wrote: "There is still no viable scientific plan to solve the risks of a recursively self-improving AI."
Drake Thomas, another Anthropic collaborator, expressed respect for Coxon's decision not to participate in the development of these models if he considers them a planetary-scale risk. Thomas commented: "Things are moving too fast, and we don't even have the degree of certainty we would like for an artificial superintelligence."
The allegations generated a reaction from Elon Musk, who adopted a significantly more skeptical stance regarding the warnings. Musk wrote on X: "This looks like a setup." He later suggested that the incident could be part of a 'psyop,' a term used to describe a psychological operation aimed at shaping public perception.
Musk responded to a post by Parker Thayer, a researcher at the conservative group Capital Research, who raised, without providing evidence, the theory that Coxon's statement would be the beginning of a sophisticated, funded public relations operation to gain support for stricter regulations against AI. Musk claimed that preparations for this supposed operation had been underway for some time and that Coxon's publication was merely the initial trigger of the process.
Coxon directly rebutted the billionaire, posting a selfie and asserting that he was a real person and that those were his genuine convictions. He also mentioned that Musk could consult him with xAI researchers if he hadn't fired some of them.
More information:
In response to the doubts, Anthropic defended its strategic approach. A company spokesperson informed The Guardian: "We have always been transparent that AI will bring enormous benefits and unprecedented risks." The company emphasized that its models continue to be developed with some of the industry's most robust safety measures. However, not all experts agree that the primary risk should be a superintelligence capable of annihilating humanity.
Gary Marcus
Gary Marcus, a well-known figure in the AI debate, opined that the time has come to boycott the technology due to the damage already occurring. For Marcus, the most urgent concerns include AI-assisted pathogens, conflicts initiated or intensified by disinformation generated by AI systems, and attacks capable of destroying vital infrastructure. He declared: "I haven't seen anything to indicate that any of these things are under control."
The discussion took on a particularly concrete aspect on the same day the researchers' warnings gained momentum. Anthropic released a report detailing how the company dismantled an attempt to build a biological weapon using its AI models. This event does not prove that an autonomous AI developed a weapon on its own, but it illustrates how currently available models can be employed in attempts to use for extremely dangerous purposes. The case placed side-by-side two usually separate debates: the future risks of potentially superintelligent AI and the more immediate dangers related to the use of current models for activities that can cause serious harm. The original article was published in Olhar Digital.
