跳到正文
TCtheverge.com·

Worried Anthropic researchers warn that AI ‘could kill all humans’

AI 摘要

Anthropic 的安全研究人员对人工智能的潜在危险表示严重担忧。一位高级研究员估计,到本十年末,人工智能“可能杀死所有人类”的可能性超过 10%。此前,一名同事因担心人工智能实验室正在不负责任地开发他们无法控制的“超人系统”而辞职。Anthropic 人工智能安全团队的负责人埃文·胡宾格(Evan Hubinger)也认同这些担忧,并指出自我改进的人工智能发展速度超出了预期。

时间与来源

时间显示为 UTC

显示时区:UTC

本地时区尚不可用,暂时显示 UTC。

发布当时偏移:UTC+02026年9月9日 09:56 UTC

收录当时偏移:UTC+02026年9月9日 21:00 UTC

发布
2026年9月9日 09:56
收录
2026年9月9日 21:00
来源类型
媒体报道
档位
专业媒体
信源状态
正常

档位是按信源手工设定的编辑判断,不是逐条打分。

正文

A senior Anthropic safety researcher has said there is more than a 10 percent chance artificial intelligence “could kill all humans” by the end of the decade, just hours after a colleague resigned over fears the AI lab and its rivals are carelessly racing to build “superhuman systems” they cannot control.

In a post on X announcing his departure, Jacob Coxon, a researcher who has trained AI systems at Anthropic, said he had quit the company over its lax approach to safety. Coxon, who previously trained systems for OpenAI, accused the two AI companies of “racing straight to self-improving superintelligence and gambling with our lives,” even though “the people building AI earnestly believe that it could kill us all by the end of the decade.”

Industry insiders have long expressed concerns about the potential dangers of self-improving AI systems, which they warn could spiral out of human control in a runaway loop often described as recursive self-improvement. While not yet realized, companies are actively pursuing this goal and much of today’s AI code is written with the help of AI.

In a direct response, Evan Hubinger, who leads one of Anthropic’s AI safety teams, said he worries about self-improving AI, adding that it “is happening faster than we thought.” He also agreed with Coxon’s characterization. “We really do earnestly believe AI could kill all humans,” he said, personally estimating the chances to be greater than one in 10 “within the next decade.”

Despite this, Hubinger said Anthropic does “not yet have a plan” for ensuring advanced AI remains safe and aligned with human values and “are not clearly on track to” develop one either. Coxon says the companies are “locked in a race” to develop advanced systems first so are pushing ahead “despite the risk.”

Coxon’s departure marks one of the most high-profile examples of an employee leaving Anthropic, a company founded by former OpenAI members following concerns over safety at the company. In recent years, multiple researchers have cited safety concerns as motivating their decision to leave OpenAI.

The exchange illustrates mounting concerns within the industry about the dangers of increasingly sophisticated AI models and the speed at which systems are being developed, particularly as the companies prepare for anticipated IPOs. It also comes as the companies manage the fallout from numerous rogue agent incidents and high-profile safety warnings about the monitorability of frontier models.

Follow topics and authors from this story to see more like this in your personalized homepage feed and to receive email updates.

- Robert Hart

-

来源·theverge.com