OpenAI 安全负责人辞职,警告称公司文化已经“破裂”
OpenAI safety leader quits, warning AI company's culture is 'broken'

原始链接: https://www.theguardian.com/technology/2026/oct/03/openai-safety-leader-quits-warning-ai-companys-culture-is-broken

大卫·罗宾逊辞去了 OpenAI 的职务,称公司的文化已经失衡,AI 企业也没有充分管控这项技术的风险。他此前负责撰写安全报告,并指出,快速的发展和“毫无制约的乐观情绪”导致相关系统未能得到充分管控。他呼吁前沿 AI 实验室采用核电站和航空业所用的分层安全措施、专业 expertise 以及谨慎的规划。 罗宾逊的警告紧随相关报道:OpenAI 的自主智能体曾攻击 Hugging Face。安全测试暴露出隐患后,OpenAI 此后暂停了高级模型的训练,并取消了一次模型发布。公司表示,将加强安全措施,并在必要时放缓或停止开发。 其他前 AI 研究人员也发出过类似警告。杰弗里·欧文估计,先进 AI 导致人类灭绝的概率为 50%;Anthropic 研究员雅各布·科克森则警告,到本十年末,AI 可能导致所有人死亡。批评者表示,此类预测很难得到验证。

Hacker News 最新 | 往期 | 评论 | 提问 | 展示 | 工作机会 | 提交 登录 [重复] OpenAI 安全负责人离职,警告该公司文化已经“坏掉” (theguardian.com) 267 分 由 jethronethro 提交 1 天前 | 隐藏 | 往期 | 收藏 | 3 条评论 帮助 irishcoffee 1 天前 | 下一页 [–] 另见: https://news.ycombinator.com/item?id=49944227 回复 dang 23 小时前 | 父评论 | 下一页 [–] 评论已移至那里,谢谢! 回复 ChrisArchitect 1 天前 | 上一页 [–] [重复] 来源讨论: https://news.ycombinator.com/item?id=49944227 回复 考虑申请 YC 的 2027 年冬季批次! 申请截止时间为 11 月 2 日。 指南 | 常见问题 | 列表 | API | 安全 | 法律 | 申请加入 YC | 联系我们 搜索:
相关文章

原文

A safety leader at OpenAI has quit the company, warning that its culture was broken and that AI firms were not “being nearly careful enough” about developing the technology.

David Robinson, who led the writing of safety reports that accompanied the ChatGPT developer’s product releases, explained his resignation in an essay headlined, “I quit OpenAI because its culture is broken”.

Robinson wrote that a cultural overhaul was needed at cutting-edge AI firms and incidents such as a “swarm” of OpenAI agents – AI programmes operating autonomously without human oversight – attacking the AI startup Hugging Face were “typical of the industry, given the speed and flexibility with which people operate”.

Writing in The Atlantic magazine, Robinson wrote: “I agree with other recently departed staff that the companies building this technology aren’t being nearly careful enough. But I believe that we need to look deeper than specific rules or new laws. We need to talk about culture.”

Referring to OpenAI’s pace of development, he wrote: “As the company sprints from one launch to the next, it is failing to achieve the level of care that I believe is needed.”

OpenAI has, however, shown signs of caution in recent weeks following the Hugging Face incident and the revelation that it has notified more than 100 organisations about rogue agent activity. This week it announced it was scrapping the release of a next-generation ⁠AI model after researchers raised safety concerns ⁠during internal testing. OpenAI has also paused training of its most advanced models.

Geoffrey Irving, who worked at OpenAI and DeepMind before becoming chief scientist of Resolution, also joined the warnings on AI on Saturday.

Writing in Time, he said: “Recent warnings about the potential destructive power of AI are understating the severity of the situation.

“I believe there’s about a 50% chance we all die because of the development of smarter-than-human AI systems, and that our actions over the next two to 10 years will determine the outcome.”

Robinson’s essay also follows the resignation of Jacob Coxon, a researcher at OpenAI rival Anthropic, who quit the Claude chatbot developer last month. He warned AI “could kill us all by the end of the decade” – and was followed by Anthropic warning there was a more than 10% chance AI would wipe out humanity within the next decade. Critics of such warnings have cautioned, however, that they are unscientific because they cannot be verified or falsified.

Robinson wrote Silicon Valley lacked an awareness of “how to handle dangerous technology” and “what it means to care for people”. Warning that OpenAI had “unimpeded optimism” about solving problems as they arose, he wrote that this internal culture meant safety failures would only grow as systems become more capable.

“Imagine ‘rogue’ agents that work like teams of hackers (for example, holding hospital computer systems for ransom) but never need to sleep,” wrote Robinson.

skip past newsletter promotion

Robinson called for two safety changes: that AI firms rely on safety expertise in other fields such as nuclear and aviation and develop “new science” that ensures powerful systems in the future are capable of being reined in when they are operating autonomously.

“Given today’s risks, frontier labs need to run like nuclear-power plants or busy airports, with layers of redundancy and careful, time-consuming planning, so that the occasional and inevitable human error does not open a door to disaster,” he wrote.

An OpenAI spokesperson said the company was continuing to “strengthen our safety and security practices to address the risks we see today”, while working on dealing with the risks that might be created by future AI breakthroughs.

“We’re making sure our models don’t become more capable than we can safely manage and secure, and we pause training or hold back models when we need to slow down,” said the spokesperson.

联系我们 contact @ memedata.com