What an Ex-Anthropic Researcher’s Warning About Human Extinction Really Means

News Room
12 Min Read

A string of posts on X from an AI researcher who says he quit a job at Anthropic over ethical concerns has gone viral, prompting thousands of responses and reigniting debate over the potential future dangers of artificial intelligence.

Jacob Coxon, an AI researcher who also worked at OpenAI on pretraining AI models, resigned from Anthropic over concerns that AI development could lead to human extinction. “Neither company is acting responsibly,” he wrote. “They are racing to self-improving superintelligence and gambling with our lives.” Superintelligence refers to a hypothetical scenario in which AI’s “cognitive” abilities vastly exceed those of humans across every domain — science, math, strategy, creativity and problem-solving.

The threaded post on X from Coxon has more than 156 million views as of this writing and has been liked more than 752,000 times.

The statement quickly spread on social media and was featured on major news sites. It even drew a comment from Illinois Governor JB Pritzker, who responded on X, “It’s becoming more clear the threat AI poses to humanity, so I’m calling for immediate action from the industry and Washington.”

CNET

In less than 24 hours, the statement also drew responses from others at Anthropic who, rather than debating with Coxon, confirmed that the threat that AI will overtake and potentially eliminate humans is not a wild fantasy but a reality for AI researchers.

“Jacob [Coxon] is correct here,” wrote Evan Hubinger, a team leader at Anthropic in AI alignment stress testing, on X. Hubinger wrote in his response, “We really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade.”

Hubinger said, “I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to.”

Representatives for Anthropic and OpenAI didn’t immediately respond to requests for comment.

Resignations that act as warning signs

It’s difficult to dismiss Coxon’s concerns for one big reason: He’s not the first researcher at a major AI company to leave for these reasons. In February, a safety lead at Anthropic, Mrinank Sharma, resigned from the company, telling followers on X, “The world is in peril… We appear to be approaching a threshold where our wisdom must grow in equal measure to our capacity to affect the world, lest we face the consequences.”

In the same month, OpenAI researcher Zoë Hitzig wrote in a New York Times column that she quit the company in part because it was following Meta’s path of putting profits ahead of ethics. Days later, she reposted a poem she’d written on X, titled “How We Programmed the Apocalypse.”

Since then, AI leaders such as OpenAI CEO Sam Altman have walked back some of their previous comments about AI’s potential for harm or whether it will create mass unemployment. Anthropic publicly acknowledged earlier this year that it has loosened some of its safety standards to keep up with competitors, including OpenAI and a swath of Chinese companies that have been releasing cheaper, high-performing AI models.

That competition, with both OpenAI and Anthropic hurtling toward initial public offerings, has raised concerns that, instead of moving cautiously in deploying AI models, industry leaders are rushing into the unknown and releasing models that could pose present and future dangers. This was the case when Anthropic released advanced AI models, which ended up being pulled or temporarily restricted due to major cybersecurity risks.

In the case of Coxon, whose resignation has caused the largest stir yet among those who’ve left AI companies, the dangers of contributing to the technology’s accelerating pace outweighed any financial gains. “I left before any of my equity vested,” he told the news site Axios.

Ironically, both OpenAI and Anthropic have taken the stance that if they slow down, someone less concerned with safety will drive AI’s future.

“Coxon is right that competition is driving the speed,” said Shama Hyder, a professor of AI at Link School of Business. “Every one of these companies believes slowing down hands the lead to someone less careful, so even the people who understand the risk have a reason to accelerate. “

AGI and superhuman AI

How quickly AI poses an imminent threat to humanity may depend on when the technology reaches two purported benchmarks that have become increasingly blurred: superhuman AI (also known as artificial superintelligence or ASI), which Coxon mentioned in his viral post, and AGI, or artificial general intelligence.

AGI would mark the hypothetical point at which AI meets (or exceeds) the capabilities of a human mind. Some predict that AI will achieve that benchmark in just a few years, or at least sometime before the end of the 21st century. That timeline has also been thrown into doubt by some AI leaders, such as Nvidia’s Jensen Huang, who downplayed the importance of such markers on a recent earnings call where he said, “For many tasks, we could say that we have already achieved AGI.” The milestones, he said, “are kind of senseless at this point.”

The goalposts are also varied. Some may choose to measure AGI using specific tests, while others would consider AGI achieved only if AI could act autonomously in ways that surpass human capabilities.

ASI would be further down the road, after AGI, and would theoretically mark the point at which AI could outperform all humans across all tasks, particularly in deep thinking and cognitive understanding. The timeline for that is much fuzzier, but the speed at which AI is advancing has researchers such as Coxon suggesting it’s sooner than many people think.

Sam Altman and Howard Lutnick on stage
OpenAI CEO Sam Altman (left) and US Commerce Secretary Howard Lutnick (right) on stage at the G20 forum this month. Altman has acknowledged the risks of AI while trying to sway the public on the technology’s benefits.Katelyn Chedraoui/CNET

A recent event that may have some adjusting their timeframes is the incident exposed in July in which OpenAI agents went rogue, plotted together and hacked Hugging Face, an AI community that was acquired by Nvidia. The speed at which the Hugging Face incident played out and the likelihood that it could happen again — this time without warning — have fueled warnings about the dangers of autonomous AI systems.

Doomers vs. boosters

The debate over whether or not AI will kill all humans by 2030 mirrors the broader discussion over AI doomers versus boosters, and the lack of nuance in the middle. While AI boosters typically hype AI by emphasizing its extraordinary potential, such as huge productivity gains, AI doomers hype it in the opposite direction, emphasizing job loss and human extinction.

As Emily Bender and Alex Hanna argue in their book The AI Con, both AI boosters and doomers rely on science-fiction tropes to frame superintelligence as either a utopian savior or an apocalyptic threat, despite a lack of evidence that such technology is near. Many AI critics argue that either narrative functions as corporate marketing, with the same underlying premise, falsely presenting an all-powerful AI future as an absolute inevitability.

The speed at which Coxon’s post went viral and the response it’s drawing may be an indication that the AI industry has another PR crisis on its hands in a year when AI data centers are under fire and trying to change their image, and social media companies, including Meta, are reeling from multibillion-dollar judgments over issues like tech addiction.

One approach that AI leaders like Altman have taken is to focus the public’s attention on what AI brings to the table and what it could offer beyond job losses and mass death — in other words, swapping doomer rhetoric for booster rhetoric.

“Abundance” has been a buzzword that’s been thrown around of late, promising a world in which AI’s advances and benefits will eliminate scarcity and improve life for everybody by lowering barriers to creating new goods and services and providing them more cheaply, and not just for those who can afford the costs of advanced AI technology.

Hyder said that what was missing from Coxon’s post was guidance on what people outside the AI industry can do. “The extinction debate belongs to a few hundred people who control training runs and the regulators who can inspect them,” she said.

The question ultimately boils down to what humans want to do with the technology. “If AI is powerful enough to do the dramatic things people fear, it’s powerful enough to do the dramatic things we’ve been hoping for in medicine and science,” Hyder said.

So far, the hyped-up abundance argument has been met with skepticism, especially when it keeps public attention and fascination focused on AI. Boosterism benefits the companies that are guzzling investment dollars and promoting the billions (or trillions) their companies’ IPOs may represent.

Who cares about abundance, you could argue, if there are no humans left in the world to enjoy it?

Read the full article here

Share This Article
Leave a comment

Leave a Reply

Your email address will not be published. Required fields are marked *