跳到正文
@rohanpaul_ai· @rohanpaul_ai · X·· 27 天前AI 评分44
AI 导读

Anthropic 研究员 Jacob Coxon 因担忧递归自我改进导致模型失控而离职,他警告最快明年底局势可能已失控。他为此从 OpenAI 转投 Anthropic 从事安全研究,却认为即便有真诚的防护措施,也无法在缺乏协同约束的竞争下奏效。Anthropic 2 万亿美元 IPO 凸显矛盾:既要市场相信强大 AI 能支撑巨额估值,又承认能力越过危险阈值时或需放缓开发。

正文

WSJ just reported: Anthropic researcher Jacob Coxon is quitting AI because he thinks self-improving models could become uncontrollable by 2027.

“We’re on track for a lot of the most aggressive of these scenarios where by the end of next year things could be out of control already,” he said.

His fear centers on recursive self-improvement, where AI takes over enough AI research to speed development of increasingly capable successors.

He moved from OpenAI to Anthropic specifically for its safety work, yet says even sincere safeguards cannot overcome competition without coordinated restraint.

The most interesting point is, Coxon is leaving despite believing Anthropic takes safety seriously.

Anthropic's $2 tn IPO makes that tension so obvious. Anthropic is simultaneously asking the world to believe 2 things: that increasingly powerful AI can create enormous economic value, potentially supporting a $2 trillion public valuation, and that development may eventually need to be slowed when capability crosses dangerous thresholds.

Those positions create a difficult governance test: will safety commitments still bind when obeying them becomes commercially expensive.

来源:@rohanpaul_ai · x.com