Rohan Paul· @rohanpaul_ai · X·· 3 小时前AI 评分61
AI 导读
据 WSJ 报道,三名因涉嫌不当行为被 OpenAI 解雇的安全研究员已要求其董事会叫停任何进一步削弱人类监控 AI 推理能力的开发。其中 Tomek Korbak 和 Mikita Balesni 曾主导 2025 年一篇论文,警告链式推理监控虽有用但脆弱。OpenAI 则表示解雇涉及敏感信息处理不当。
正文
WSJ: Three safety researchers fired by OpenAI for alleged misconduct have asked its board to halt any development that further reduces humans’ ability to monitor AI reasoning.
Two of them, Tomek Korbak and Mikita Balesni, led the 2025 paper that warned chain-of-thought monitoring is useful but fragile.
OpenAI however says the dismissals concerned mishandled sensitive information.
来源:Rohan Paul · x.com