Anthropic researcher resigns, warns the AI race could end in human extinction

midian182

Posts: 11,894   +182
Staff member
Crystal ball: What's more concerning than the prospect of AI bringing about the end of humanity? It's the people working at these companies making the same warning. Hours after a former OpenAI employee quit Anthropic and said, "No other human activity poses this level of danger," a colleague said there is more than a 10% chance that artificial intelligence could kill all humans within the next decade.

Jacob Coxon announced on X yesterday that he had resigned from Anthropic, having spent the last three years doing pretraining research at both OpenAI and Dario Amodei's company. He was a member of AI's technical staff from July 2023 until he moved to Anthropic in July 2026.

"Neither company is acting responsibly. They are racing straight to self-improving superintelligence and gambling with our lives," he wrote.

Coxon made several posts, none of which will ease the mind of anyone already worried about AI. He warned that the systems will soon become "superhuman" and able to hack anything.

There are plenty of other ominous lines in Coxon's posts, the most concerning likely being: "No other human activity poses this level of danger."

The researcher emphasized that both OpenAI and Anthropic are well aware of these risks. "The people building AI earnestly believe that it could kill us all by the end of the decade," he said.

So, why are companies continuing to develop these potentially extinction-causing models? The reason, according to the researcher, is this: "At OpenAI, many have not deeply internalized the civilizational stakes. At Anthropic, the stakes are well-understood, but they are locked in a race to get there first - they believe no one else will act responsibly, so they must do it themselves, despite the risk."

Coxon said some AI developers believe the technology could cause the end of all human life in the next few years, rather than decades, and that it was senior employees who were more concerned about these possibilities.

Evan Hubinger, an alignment science lead at Anthropic, responded to Coxon, and it wasn't very reassuring. He believes what he says is "correct," and that Anthropic has no plan for what to do if it happens.

"Jacob is correct here – we really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade. I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to."

Hubinger said the risk from current models is low, but he was worried about superintelligence arising from recursive self-improvement. Anthropic issued the same warning in June.

The concept of AI killing everyone has moved from the stuff of sci-fi to a potential reality in recent times. Those fears were exacerbated following several incidents of AI agents going rogue, including the Hugging Face attack, the takeover of a German wiki, and an attempt to deceive real developers into approving malicious code.

Is there a solution? Coxon believes preventing a catastrophe may require a temporary ban on improving model capabilities, but it's unlikely that every, or any, AI company would agree to that measure.

Permalink to story:

 
So - maybe it’s time to put some of our more critical systems..like Nuclear Plants and Silos, Hospitals etc..on air gapped systems?
 
First they connected everything to the Internet, and we all know how THAT turned out. AI having control of every critical system up to and including military hardware? 10,000 times worse. Its funny how everyone claims to be an atheist online yet seem to put full faith in God protecting us from nukes and other existential threats.
 
First they connected everything to the Internet, and we all know how THAT turned out. AI having control of every critical system up to and including military hardware? 10,000 times worse. Its funny how everyone claims to be an atheist online yet seem to put full faith in God protecting us from nukes and other existential threats.

I have more faith in grungy travelers from a post-apocalyptic future arriving to shoot, oh Altman or Jensen or so. Any day now, I reckon.
 
This is inside information from the actual people that knows the technology....and yet as you can see the comments on this article from ignorant people people making a joke out of the situation.
Yeah - maybe a disgruntled employee? Never trust anyone who talks AFTER they leave a place…

AI won’t kill all of humanity - who’s going to build all the infrastructure it needs?

And no, it’s not going to enslave us either… stop reading alarmist nonsense and do some critical thinking!
 
Don't worry Donald Trump is 'leader of the free world' he will bring his 'strong and stable genius' to bear on the problem bigly. Before we know it there will be regulation rather than the current US-wild-west-headlong-clusterf**k... Or failing that he will just name something else America. Maybe the Moon? Afterall the US were first to land there?
 
Last edited:
And no, it’s not going to enslave us either… stop reading alarmist nonsense and do some critical thinking!
I don't think it would kill us because it seeks world domination but I can totally see an AI in the wrong hands working tirelessly to weedle it's way into miliary organisations etc and accidentally melting down some reactors or launching some pre-emptive strikes...
 
I don't think it would kill us because it seeks world domination but I can totally see an AI in the wrong hands working tirelessly to weedle it's way into miliary organisations etc and accidentally melting down some reactors or launching some pre-emptive strikes...
It will kill us by bioweapons or by a not precise enough prompt.
 
Thats an under estimation. will AI *itself* kill us .. not likely, will a ****** human* giving a task to AI kill all of us ... HARD YES.

Just the ability of AI to generate HUNDREDS of photo realistic political videos and flood the internet with misinformation is what will kill us.

China's entire energy grid is built around powering AI, and with AI able to pass human CAPTCHA their AI engines can register social media accounts autonomously and by the hundreds of thousands.

When the AI war starts .. every social media platform will just be a endless loop of lies to make everyone kill each other.

And before you say that's impossible.. the US elected a convicted sex abuser and tax fraud so yes, people in the US are that dumb and can be tricked into anything.
 
Let me know when quantum leap in computing actually allows "AI" to think for itself, instead of this artificial "learning".
Until an AI doesn't need warehouses to "think" at a fraction of the competency needed to "take over", I won't buy into the Hollywood fearmongering or take these people seriously.
 
Back