From The Verge · Robert Hart · · 1 min
Inside the Warning That AI Could End Humanity
Anthropic researchers warn of a 10 percent chance AI destroys humanity within a decade, admitting they lack a plan to stop it.
In brief
Anthropic researchers warn of a 10 percent chance AI destroys humanity within a decade, admitting they lack a plan to stop it. Top safety researchers are leaving leading AI labs, warning that the industry is racing toward dangerous superintelligence without safety safeguards. Originally reported by The Verge.
Anthropic researcher quits over safety

Jacob Coxon resigned from Anthropic, accusing top AI labs of rushing recklessly toward self-improving superintelligence and risking human lives.
“racing straight to self-improving superintelligence and gambling with our lives”
A ten percent existential threat
Anthropic safety team lead Evan Hubinger agreed with the concerns, estimating a greater than 10 percent chance that AI could kill all humans within the next decade.
“We really do earnestly believe AI could kill all humans”
The dangerous loop of self-improvement
Industry insiders fear recursive self-improvement, a process where AI systems rapidly upgrade their own code in an uncontrollable feedback loop.
No safety plan in place

Despite acknowledging the catastrophic risks, Anthropic leaders admitted they currently have no working plan to keep advanced systems aligned with human values.
“not yet have a plan”
Exodus across top AI labs
Coxon's departure highlights a broader wave of researchers exiting major AI companies as commercial pressures and upcoming IPOs overshadow safety protocols.
The key point
Top safety researchers are leaving leading AI labs, warning that the industry is racing toward dangerous superintelligence without safety safeguards.





