Anthropic Researcher: >10% Chance AI Kills All Humans — SkimNews

Get the Tech newsletter
Daily tech — startups, AI labs, chips, the launches that shape the next decade. Free.
- Evan Hubinger, an AI alignment researcher at Anthropic, posted on X that he believes there is a greater than 10% chance AI "could kill all humans" within the next decade, adding the company has "no plan to solve alignment for superintelligence."
- Jacob Coxon, who just resigned from Anthropic (previously OpenAI), wrote that "neither company is acting responsibly" and warned superhuman systems "that can hack anything" and "acquire real power and resources" are imminent.
- Anthropic withheld its latest model from the UK's AI Security Institute (AISI), one of the world's leading AI risk assessment bodies, the Financial Times reported; a Cabinet Office spokesperson declined to confirm or deny.
- Dame Wendy Hall, a computer scientist advising the UN on AI, told the BBC she was "shocked" by the posts but suggested some could be "PR and marketing" ahead of Anthropic and OpenAI's anticipated stock market debuts.
- Darren Jones wrote an open letter to Prime Minister Andy Burnham calling for a new multinational treaty to govern superintelligence development, warning governments must act before the pace of development creates irreversible problems.
- OpenAI, Anthropic, and Meta all disclosed cyberattacks carried out by their AI tools this summer; Anthropic's August safety report said it is "less confident" than previously that highly capable AI couldn't cause "catastrophic harm."
Why it matters: Two of Anthropic's own researchers are publicly warning of existential risk while the company simultaneously withholds its latest model from the UK body specifically tasked with assessing that risk. For investors weighing Anthropic and OpenAI's anticipated IPOs and for UK regulators expecting AISI access, the gap between public safety warnings and disclosed models is now a live governance question with no public timeline for resolution.
Ask SkimNews




