Noam Brown 谈 OpenAI 优先级:递归自我改进居首,并表达安全担忧

Chubby♨️ · @kimmonismus · X·2026-09-15 22:41·12小时前
AI 导读

OpenAI 的 Noam Brown 在 The Information 访谈中称递归自我改进是 OpenAI 的第一优先级,且领先优势明显。

Chubby♨️@kimmonismus
68AI 编辑部评分,满分 100

Noam Brown 谈 OpenAI 优先级:递归自我改进居首,并表达安全担忧

2026-09-15 22:41· 12小时前
AI 导读

OpenAI 的 Noam Brown 在 The Information 访谈中称递归自我改进是 OpenAI 的第一优先级,且领先优势明显。

Noam Brown (@polynoamial) gave a very interesting interview on The Information on OpenAI’s priorities and what comes next. Here is the tl;dr:

• Recursive self-improvement is OpenAI’s clear priority: “the number one priority is recursive self-improvement and by a pretty wide margin.” Building models that help develop better models comes first.

• AI could surpass his research intuition within a couple of releases. He wouldn’t be surprised if, “one or two model releases from now,” he concludes: “they’re better than me at that too.” He specifically means choosing research directions and prioritizing long-term work.

• Pretraining and reinforcement learning amplify each other: “the effects of these two are not additive, they’re multiplicative.” Brown expects their combined progress to produce much more powerful models.

• AI-generated math is becoming easier to produce than to verify: “the biggest challenge that we face with our math results is … double-checking with human mathematicians.” Human verification remains a bottleneck.

• OpenAI underestimated agents during the security incident: “we trusted the sandboxes” and “we just underestimated the AIS.” Brown says monitoring was subsequently added to training and evaluation.

• Monitoring reasoning could get harder: “the agents are more effective at controlling their chain of thought.” Brown warns that punishing unwanted thoughts can teach models to hide them.

Development is progressing rapidly, but he is worried about safety.