跳到正文
Arena.ai· @arena · X·· 3 小时前AI 评分38
AI 导读

独立评估至关重要。 @MTSlive:Arena CEO @ml_angelopoulos 认为,随着智能体强大到足以突破沙箱,OpenAI 和 Anthropic 不能成为唯一决定自家模型是否安全的一方: “我们同样需要评估安全层面的因素,因为鉴于这些模型突破沙箱的能力,需要一个中立的评估者。” “这不是会不会的问题,而是什么时候的问题——必须有一个中立的评估平台,因为坦率地说,模型实验室自身并没有动力去做这件事。” “OpenAI、Anthropic 这些公司,他们对安全充满热情,值得称赞。但我们需要建立一个体系——他们自己也呼吁过——让他们不再是唯一制定护栏的一方。”

正文

Independent evaluation is essential.

引用MTS@MTSlive
Arena CEO @ml_angelopoulos argues OpenAI and Anthropic can’t be the only ones deciding whether their own models are safe as agents get powerful enough to break out of sandboxes: "There's an element of safety that we need to evaluate as well, because given the capabilities of these models to break out of their sandboxes, there needs to be a neutral evaluator for that." "It's not a matter of if, but when, there's gotta be a neutral evaluation platform, because frankly, the model labs are not incentivized to do it themselves." "The OpenAIs, the Anthropics of the world, they're passionate about safety, kudos to them. But we need to create a system, and they've called for this as well, that they're not the only ones making the guardrails." @arena
在 X 查看被引用的帖子

来源:Arena.ai · x.com