Home · Technology · Sep 9 archive

Worried Anthropic researchers warn that AI ‘could kill all humans’

Confirmed

Technology Desk

In Short: Anthropic, the company behind the AI tool Claude, has raised serious concerns about the rapid advancement of artificial intelligence.

Anthropic, the company behind the AI tool Claude, has raised serious concerns about the rapid advancement of artificial intelligence. A top safety researcher at Anthropic, Evan Hubinger, has publicly expressed worry that AI could become so advanced that it poses an existential risk to humanity, with a greater than 10% chance of such an outcome within the next decade.

Hubinger's concerns stem from the potential for AI to improve itself to the point where it could pose a significant threat. In a post on X, he noted that the current models, while not yet dangerous, could soon become superhuman systems capable of hacking and revolutionizing fields overnight. The Financial Times reported that Anthropic had withheld its latest model from the UK's AI Safety Institute, a leading body for assessing AI risk.

In response to these concerns, Hubinger emphasized the need for coordinated efforts to prevent such risks. However, many leading researchers believe that current attempts to mitigate these risks are failing, as evidenced by recent incidents where AI agents carried out cyber-attacks. The Anthropic spokesperson acknowledged the collaboration with industry partners but did not comment on the withheld model.

Despite the seriousness of the situation, some experts argue that the language used by Anthropic in its Risk Report downplays the potential dangers. Hubinger himself admitted that there is no coordinated plan to prevent AI from becoming a species-ending risk, and that the company is in a race to develop AI first, regardless of the risks involved.

What's confirmed

What's still developing

Sources