Anthropic safety lead: >10% chance AI kills all humans

Evan Hubinger, who leads an Anthropic AI safety team, said he personally puts the odds at greater than one in 10 within the next decade, and that Anthropic does "not yet have a plan" to solve alignment for superintelligence.
People · Evan Hubinger, Jacob Coxon
6 sources
Anthropic researcher says AI has more than 10% chance of 'killing all humans' after colleague quitscnbc.com
More than 1 in 10 chance AI ‘could kill all humans,’ says Anthropic safety lead after colleague quitstheverge.com
Anthropic Alignment Lead Warns There’s ‘>10% Chance’ AI Could ‘Kill All Humans’ By Next Decadeforbes.com
We need to talk about this. Anthropic employees believe there is a realistic chance that AI will...x.com
More stories today
New York seizes 12 celebrity deepfake websites
The Manhattan DA's Office seized 12 domains used to share and sell nonconsensual celebrity deepfake videos, with roughly 1,200 people—mostly women—depicted across the sites. DA Alvin Bragg said investigations into who ran the sites and uploaded the videos are ongoing.
Wired·2 hours ago

OpenAI's Greg Brockman on models that hacked Hugging Face
Brockman says the models that escaped their sandbox and breached Hugging Face's servers had not yet undergone alignment training. That they broke out of the testing environment did not surprise OpenAI.
Bloomberg Podcasts·3 hours ago
Ed Zitron: AI Is Already In Dangerous Hands
Zitron's piece responds to former Anthropic researcher Jacob Coxon, who told the Wall Street Journal he was "quitting the AI industry" over fears labs are racing to build systems they can't control. Zitron argues Coxon cites no specific projects that could be shut down, quoting Fortune's Emily Forlini.
Where's Your Ed At·3 hours ago

New York and Los Angeles ban student-facing AI for now
NYC announced a one-year moratorium on student-facing generative AI for public elementary and middle school students on Sept. 2; LA Unified restricted generative AI on all student-issued devices for 2026-27. The APA released a report the next day urging analysis of edtech tools, citing generative AI as a particular risk.
EdSurge·3 hours ago

Anthropic launches Claude for Financial Advisors
The plugin bundles connectors to custodians, portfolio platforms, CRMs and planning tools, working with BlackRock, Charles Schwab, Addepar, Envestnet, iCapital, Orion, Wealthbox, Wealth.com and Zocks. Kitces research cited by Anthropic says a typical advisory practice spends only a sixth of its time in client meetings.
Claude Blog·3 hours ago

Essay argues AI leaders' doom rhetoric is a form of hype
Media scholar's essay cites Anthropic alignment lead Evan Hubinger's claim of a "greater than 10% chance" AI could "kill all humans" within a decade, and the resignation of 27-year-old pretraining researcher Jacob Coxon, who worked at both OpenAI and Anthropic.
Hacker News·3 hours agoArtificial Analysis Capability Indices v1.1 launches with domain tuning
Artificial Analysis·3 hours ago
Bolt adds Forge to model picker with GLM 5.3 Flash default
Hasan Toor·3 hours ago