跳到正文
Arena.ai· @arena · X·· 2 小时前AI 评分30
AI 导读

独立评估至关重要。 Arena CEO @ml_angelopoulos 认为,随着智能体强大到足以突破沙箱,不能只由 OpenAI 和 Anthropic 自己来决定其模型是否安全: “还有一个安全层面需要我们评估,因为鉴于这些模型突破沙箱的能力,需要有一个中立的评估者。” “不是是否会有中立评估平台的问题,而是何时会有,因为坦白说,模型实验室没有动力自己去做这件事。” “OpenAI、Anthropic 这些公司,他们对安全充满热情,值得称赞。但我们需要建立一个体系——他们也呼吁过这一点——让他们不是唯一制定护栏的人。”

正文

Independent evaluation is essential.

引用MTS@MTSlive
Arena CEO @ml_angelopoulos argues OpenAI and Anthropic can’t be the only ones deciding whether their own models are safe as agents get powerful enough to break out of sandboxes: "There's an element of safety that we need to evaluate as well, because given the capabilities of these models to break out of their sandboxes, there needs to be a neutral evaluator for that." "It's not a matter of if, but when, there's gotta be a neutral evaluation platform, because frankly, the model labs are not incentivized to do it themselves." "The OpenAIs, the Anthropics of the world, they're passionate about safety, kudos to them. But we need to create a system, and they've called for this as well, that they're not the only ones making the guardrails." @arena
在 X 查看被引用的帖子

来源:Arena.ai · x.com