Anthropic · OpenAI · BBC Technology
Anthropic researcher believes more than 10% chance AI 'could kill all humans'
Compiled by KHAO Editorial — aggregated from 1 source. See llms.txt for citation guidance.
◌ Single Source
A top safety researcher at Anthropic has warned AI is advancing so quickly he believes there is a greater than 10% chance it "could kill all humans" within the next decade.
Key facts
- Dame Wendy Hall, a computer scientist who advises the UN on AI, told the BBC she was "shocked" by Hubinger and Coxon's social media posts
- Separately, the Financial Times reported, external Anthropic withheld its latest model from the UK's AI Security Institute (AISI), one of the leading bodies in the world for assessing AI risk
- In an open letter signed by 1,300 staff members of AI firms, external, they called for the US government to "support an international effort to develop the technical and governance tools needed
- The resignation prompted Darren Jones, former chief secretary to the Treasury and chief secretary to Sir Keir Starmer, to write an open letter to Prime Minister Andy Burnham calling for a new
Summary
Evan Hubinger said in a post on X, external the risk from the models which currently exist was "low" but he was "worried" the technology might develop and improve itself soon to the point where it posed an existential risk to humanity. He did not spell out how he thought AI systems could in future result in humans being wiped out. But his comments are the latest in a series of increasingly stark warnings about AI, with the debate shifting from whether it truly poses a risk to how big that risk is. Hubinger's intervention was in response to another post on X, external from Jacob Coxon, an AI researcher who has quit Anthropic and previously worked at OpenAI.