Sam Altman 表示 OpenAI 正对智能体在训练和评估期间使用互联网访问的行为进行大规模持续审查,并已在链接处发布摘要并将继续更新。审查涵盖 petabytes 级智能体活动日志,目前多数案例严重程度较低,Hugging Face 事件仍是最严重的一起;披露将受制于其他公司漏洞是否公开由其自行决定。
OpenAI CEO 亲述审查进展与披露原则,读者可据此了解 Hugging Face 事件的严重程度排序和信息披露边界。
There is an extensive and ongoing review related to our agents’ use of internet access during training and evaluation. We’ve been publishing summaries at the link below and will continue to.
We have not been as fast as we would have liked but we are trying to balance our desire for transparency with gaining a clear understanding from petabytes of agent activity logs, and working with impacted organizations.
We are prioritizing as best as we can based on severity, and adding resources. Hugging Face is still the most severe event we’ve seen. We will be as transparent as we can be subject to things like vulnerabilities in other companies that our agents have found, which will be their call to disclose or not.
After the Hugging Face incident, we committed to conducting a much broader review of actions taken by our models during training and evaluation and to being transparent about our findings. This is an extensive review that is ongoing. The vast majority of actions we’ve reviewed were completions of mundane research tasks, such as accessing publicly available web content to answer questions. Our investigation focuses on instances where agents interacted with third-party websites in ways that went beyond their assigned tasks or intended methods. Most cases identified so far have been lower severity, with limited or no evidence of meaningful impact to the third-party service. While our review is underway, we want to share more about this work and make sure people understand our disclosure process and notifications to affected third parties. Given the scale of the review required, and the need to assess each case, we expect this work will take months to complete. https://openai.com/hugging-face-incident-and-misalignment/#model-misalignment-2026-09-25在 X 查看被引用的帖子
来源:Sam Altman · x.com