A top safety researcher at Anthropic has warned AI is advancing so quickly he believes there is a greater than 10 per cent chance it ‘could kill all humans’ within the next decade.
Evan Hubinger said in a post on X the risk from the models which currently exist was ‘low’ but he was ‘worried’ the technology might develop and improve itself soon to the point where it posed an existential risk to humanity.
He did not spell out how he thought AI systems could in future result in humans being wiped out.
But his comments are the latest in a series of increasingly stark warnings about AI, with the debate shifting from whether it truly poses a risk to how big that risk is.
Hubinger’s intervention was in response to another post on X from Jacob Coxon, an AI researcher who has just quit Anthropic and previously worked at OpenAI.
“Neither company is acting responsibly,” he wrote.
“These will soon be superhuman systems that can hack anything, revolutionise any field overnight, and acquire real power and resources.”
OpenAI has been approached for comment.
Dame Wendy Hall, a computer scientist who advises the UN on AI, told the BBC she was ‘shocked’ by Hubinger and Coxon’s social media posts.
She told the BBC Radio that some of it could be ‘PR and marketing’ however, as Anthropic and OpenAI raced towards highly anticipated stock market debuts.
But she added: “Why would someone want to say that? I would plead with investors not to invest in this company if that is their value system.”
Separately, the Financial Times reported Anthropic withheld its latest model from the UK’s AI Safety Institute (AISI), one of the leading bodies in the world for assessing AI risk.
Anthropic has declined to comment on the posts by its employees or the situation with the AISI.
A Cabinet Office spokesperson did not comment on whether the latest model had been withheld from the AISI – instead saying it ‘continues to collaborate closely with industry partners, including Anthropic, to make models safer.’
In his post, which has been viewed more than 10 million times, Hubinger said ‘we really do earnestly believe’ AI poses a species-ending risk to humans.
“I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to,’ he added.