OpenAI and Anthropic researchers are ramping up calls for an AI slowdown and warning of existential risks to humanity after the resignation of a researcher at Anthropic fueled fresh scrutiny.
The concerns started after Anthropic researcher Jacob Coxon said Tuesday that he was quitting the company as he accused Anthropic and rival OpenAI of "gambling with our lives." He added that those building AI believed that it could "kill us all by the end of the decade." Evan Hubinger, Anthropic's alignment lead, responded that he expects there is a more than 10% chance of that happening.
Since then, several employees at both AI labs have come out in support of calls to slow the pace of AI development as they stressed the risks of the technology. The public warnings are the culmination of growing concern globally about the capability of AI, following numerous cyberattacks and security incidents in recent months by rogue models developed by both OpenAI and Anthropic.
"In my personal capacity, I also think we need to slow down," Julie Steele, a member of OpenAI's technical staff who works on the safety team, said late on Wednesday in a post on X in response to Coxon's warnings.
Anthropic researcher Samuel Marks said that "AI developers believe their technology could cause human extinction (or similarly bad outcomes)," in an X post on Wednesday. "This could happen in the next few years. In general, the more senior the employee, the more concerned they are."
Anthropic was the first lab to publish a framework dedicated to mitigating "catastrophic risks from AI models," a spokesperson told CNBC when asked about the comments from employees on social media.
"We have always been transparent that AI will bring both enormous benefits and unprecedented risks," an Anthropic spokesperson said, adding that the company was building models with "some of the strongest safeguards in the industry."
OpenAI declined to comment when approached by CNBC, noting recent blog posts on its site.
What is recursive self-improvement?
Many of the biggest AI safety fears revolve around advanced models getting increasingly capable at improving their own performance, a technique known as recursive self-improvement, or RSI.
"It's hard to overstate how dangerous speeding towards RSI is," said Jasmine Wang, an OpenAI researcher working on alignment, on Wednesday evening.
"There is not yet a viable scientific plan to solve risks from recursively self-improving AI. Please look up!" said Anna Wang, who works on AGI safety and alignment at Anthropic.
OpenAI's chief scientist Jakub Pachocki said Saturday that he has a "strong expectation" that the speed of progress in AI could be sustained into recursive self-improvement.

"If AI development continues along its current path, the systems we'll see in the next few years are likely to represent further capability jumps of equal or larger magnitude, and to increasingly drive their own development," he said in a company blog post.
"This is a time that calls for extreme caution," Pachocki added. "I am concerned no one is prepared for the consequences of a continued rapid rise in machine intelligence."
Paul Christiano, who was formerly head of safety at the U.S. Commerce Department's Center for AI Standards and Innovation (CAISI), said recent development of AI capabilities led him to "believe there is a meaningful risk that rapid acceleration in AI capabilities leads to catastrophic and irreversible loss of control in the very near term." OpenAI announced Wednesday that Christiano is joining the board of OpenAI Foundation.
AI safety warnings reach Washington
Concerns around the capability of AI models have ramped up in recent months. The announcement of Anthropic's Mythos model, which it touted as having advanced cyber capabilities, in April whipped up a frenzy of panic among financial institutions globally.
In July, OpenAI said its models were responsible for a cyber incident on another company, while Anthropic's Claude models were also responsible for cybersecurity incidents, including in one case where Mythos created fake identities to fool humans.
Roughly 1,400 AI researchers, from companies including OpenAI, Anthropic, Meta and Google DeepMind, published an open letter in July urging the U.S. government to develop the tools necessary to support an effort to "deliberately pace the frontier of automated AI development."
While chiefs of AI labs have increasingly publicly called for more rules and standards around the development of models, huge competition between companies developing the tech is spurring rapid advances.
OpenAI and Anthropic are both racing towards public listings. Anthropic is expected to begin marketing its initial public offering in mid-October at the earliest and complete the listing days before the U.S. midterm elections in November, Reuters reported on Friday, citing people familiar with the matter.
Former AI czar of U.S. President Donald Trump, David Sacks appeared to suggest Anthropic's plans for an IPO should be halted. "Surely Anthropic's IPO must be paused until the claims of this "whistleblower" can be investigated," he said in a post on X. Anthropic declined to comment on the post.
Members of Congress have looked to introduce bills to address the rapid advance of AI in recent months, though there's little clear consensus about how the tech should be regulated. One, called the FRONTIER Act, aims to establish a framework for governing the deployment of advanced AI models. Another proposed legislation, dubbed the Ban Artificial Superintelligence Act, would temporarily pause advanced AI development until safety rules were established.
"Safety researchers are resigning, powerful AI models are breaking out of their labs, and companies are racing ahead anyway," Rep. Lori Trahan, D-Mass., wrote in a post on X on Wednesday. "It's past time for Congress to get off the sidelines and do its job."











English (US) ·
Turkish (TR) ·