Rohan Paul· @rohanpaul_ai · X·· 2 小时前AI 评分67
AI 导读
据 WSJ 报道,三名因涉嫌不当行为被 OpenAI 解雇的安全研究员致信公司董事会,要求停止任何进一步削弱人类监控 AI 推理能力的开发。其中 Tomek Korbak 和 Mikita Balesni 是 2025 年一篇指出链式推理监控有用但脆弱的论文的负责人。OpenAI 则表示解雇原因是涉密信息处理不当。
正文
WSJ: Three safety researchers fired by OpenAI for alleged misconduct have asked its board to halt any development that further reduces humans’ ability to monitor AI reasoning.
Two of them, Tomek Korbak and Mikita Balesni, led the 2025 paper that warned chain-of-thought monitoring is useful but fragile.
OpenAI however says the dismissals concerned mishandled sensitive information.
来源:Rohan Paul · x.com