Anthropic 研究员 Jacob Coxon 因担忧自我改进模型失控而离开 AI 行业

Rohan Paul · @rohanpaul_ai · X·2026-09-09 08:47·50分钟前
AI 导读

据 WSJ 报道,Anthropic 研究员 Jacob Coxon 认为自我改进模型到 2027 年可能变得不可控,正退出 AI 行业。他称到明年底一些最激进的失控情形可能已经发生,其担忧核心是递归自我改进,即 AI 接手足够多的 AI 研究以加速开发更强的后继模型。

Rohan Paul@rohanpaul_ai
62AI 编辑部评分,满分 100

Anthropic 研究员 Jacob Coxon 因担忧自我改进模型失控而离开 AI 行业

2026-09-09 08:47· 50分钟前
AI 导读

据 WSJ 报道,Anthropic 研究员 Jacob Coxon 认为自我改进模型到 2027 年可能变得不可控,正退出 AI 行业。他称到明年底一些最激进的失控情形可能已经发生,其担忧核心是递归自我改进,即 AI 接手足够多的 AI 研究以加速开发更强的后继模型。

WSJ just reported: Anthropic researcher Jacob Coxon is quitting AI because he thinks self-improving models could become uncontrollable by 2027.

“We’re on track for a lot of the most aggressive of these scenarios where by the end of next year things could be out of control already,” he said.

His fear centers on recursive self-improvement, where AI takes over enough AI research to speed development of increasingly capable successors.

He moved from OpenAI to Anthropic specifically for its safety work, yet says even sincere safeguards cannot overcome competition without coordinated restraint.

The most interesting point is, Coxon is leaving despite believing Anthropic takes safety seriously.

Anthropic's $2 tn IPO makes that tension so obvious. Anthropic is simultaneously asking the world to believe 2 things: that increasingly powerful AI can create enormous economic value, potentially supporting a $2 trillion public valuation, and that development may eventually need to be slowed when capability crosses dangerous thresholds.

Those positions create a difficult governance test: will safety commitments still bind when obeying them becomes commercially expensive.

来源:Rohan Paul· x.com