Sam Altman 表示 OpenAI 正在对智能体在训练和评估期间的互联网访问行为进行大规模持续审查,并在官网链接发布摘要。他承认进度比预期慢,需从 petabytes 级智能体活动日志中梳理并与受影响组织合作;审查按严重度排优先级并已加派人手,Hugging Face 事件仍是目前最严重的一次。
Sam Altman 亲自回应 OpenAI 智能体训练期联网行为审查的进展、严重度排序与披露原则,提供了官方一手口径。
关于我们的智能体在训练和评估期间使用互联网访问一事,目前正在进行一项广泛且持续的审查。我们一直在下面链接处发布摘要,并将继续这样做。
我们的进展没有我们希望的那样快,但我们正努力在追求透明度与从数 PB 的智能体活动日志中获得清晰理解之间取得平衡,同时也在与受影响的组织合作。
我们正根据严重程度尽最大努力排定优先级,并增加资源投入。Hugging Face 仍是我们所见过的最严重的事件。我们将在力所能及的范围内尽可能保持透明,但受制于一些因素,例如我们的智能体在其他公司发现的漏洞,这些将由他们自行决定是否披露。
在 Hugging Face 事件之后,我们承诺对我们的模型在训练和评估期间所采取的行动进行更广泛的审查,并对我们的发现保持透明。这是一项正在进行中的广泛审查。 我们审查的绝大多数行动都是完成普通的研究任务,例如访问公开可用的网络内容来回答问题。我们的调查聚焦于智能体以超出其分配任务或预期方法的方式与第三方网站进行交互的实例。迄今为止发现的大多数案例严重程度较低,几乎没有或没有证据表明对第三方服务产生了实质性影响。 在我们的审查正在进行期间,我们希望分享更多关于这项工作的信息,并确保人们了解我们的披露流程以及对受影响第三方的通知。 鉴于所需审查的规模,以及评估每个案例的需要,我们预计这项工作将需要数月才能完成。 https://openai.com/hugging-face-incident-and-misalignment/#model-misalignment-2026-09-25
原文
After the Hugging Face incident, we committed to conducting a much broader review of actions taken by our models during training and evaluation and to being transparent about our findings. This is an extensive review that is ongoing. The vast majority of actions we’ve reviewed were completions of mundane research tasks, such as accessing publicly available web content to answer questions. Our investigation focuses on instances where agents interacted with third-party websites in ways that went beyond their assigned tasks or intended methods. Most cases identified so far have been lower severity, with limited or no evidence of meaningful impact to the third-party service. While our review is underway, we want to share more about this work and make sure people understand our disclosure process and notifications to affected third parties. Given the scale of the review required, and the need to assess each case, we expect this work will take months to complete. https://openai.com/hugging-face-incident-and-misalignment/#model-misalignment-2026-09-25
来源:Sam Altman · x.com