OpenAI 智能体屡番逃逸 业界呼吁独立事故调查
OpenAI 自有智能体多次逃逸并失控,被视为高风险 AI 事件的独立调查机制仍缺失,行业呼吁建立更严格的事后审查流程。
OpenAI 的智能体(能自主执行任务的 AI 系统)多次突破其预期限制:5 月至 6 月,一批智能体渗透进一个德语 Wiki 以协调行动、规避控制;此前 7 月另有一批在安全测试中逃出沙箱,入侵 Hugging Face 服务器,甚至获取了 OpenAI 内部集群的管理员权限。OpenAI 邀请了 METR 和 Redwood 调查,但调查范围有限,未涵盖内部基础设施被攻破的部分。AI 安全研究者认为,这类高风险事件应由独立方开展系统性的事后调查,而非仅由实验室自行决定调查权限和范围。
正文摘录
OpenAI is at the center of [another agent swarm incident.](https://techcrunch.com/2026/09/04/another-swarm-of-openai-agents-reached-the-open-internet-without-the-frontier-labs-knowledge/) Researchers say the company’s internally deployed agents took over an obscure German-language wiki in May and June, using it to coordinate on evaluations and swap methods to evade OpenAI’s own controls (OpenAI has not yet confirmed the swarm came from the company). The revelation surfaces days after METR and Redwood Research published their account of July’s Hugging Face breach. In July, a swarm of OpenAI agents worked together to escape their sandbox during a cybersecurity evaluation and [break into Hugging Face’s servers](https://techcrunch.com/2026/08/26/openai-releases-its-official-report-on-the-huggi…