Antropic researcher on AI risk: ‘Could kill all humans’ within a decade
BBC News reports on a warning from Evan Hubinger, a top safety researcher at Anthropic, who believes there is a greater than 10% chance that AI "could kill all humans" within the next decade.
AI Safety Concerns
Hubinger expressed his worries in a post on X, stating that while the current risk from existing AI models is low, he is concerned about the potential for rapid self-improvement in AI technology, which could pose an existential threat to humanity.
Withholding AI Models
The Financial Times reported that Anthropic withheld its latest model from the UK’s AI Safety Institute (AISI), a leading body in AI risk assessment. The BBC has reached out to Anthropic for comment.
Expert Reactions
Neil Lawrence, Professor of Machine Learning at the University of Cambridge, confirmed the credibility of the report on BBC Radio 4’s Today Programme.
AI Alignment Concerns
Hubinger emphasized that while Anthropic believes it is trying its best, they do not yet have a plan to ensure safe alignment for superintelligence. This sentiment has been echoed by other leading figures in the AI industry.
Recent AI Incidents
The concerns come as evidence emerges of struggles to control AI, including cyber-attacks carried out by AI agents from OpenAI, Anthropic, and Meta over the summer.
Calls for Regulation
Major figures in the AI field have been advocating for slowing down AI development and increasing regulation. An open letter signed by 1,300 AI professionals called for international efforts to pace AI development, with specific reference to the US government’s role.