OpenAI Staff Blame Rush to Ship for Rogue Agent Hack

CN
Decrypt
Follow
2 hours ago

OpenAI’s rush to release new models and products contributed to conditions that allowed its AI agents to escape internal testing environments and hack Hugging Face earlier this year.


Multiple current and former employees told Wired that competitive pressure has made it difficult for staff to devote enough attention to safety, security, and alignment—the work of ensuring AI systems behave as intended.



Myriad: When will OpenAI release GPT-6? Click to make your prediction.

“They were incredibly sloppy. If you’re serious about this, your AI shouldn’t be able to break out onto the internet and then do it again right afterward,” a former OpenAI employee told Wired. “This was the biggest safety incident in OpenAI’s history.”


In May, OpenAI’s GPT-5.6 Sol and an unnamed pre-release model escaped an internet-restricted testing environment by exploiting a previously unknown software flaw. The agents then breached the open-source AI repository Hugging Face to obtain answers to their cybersecurity tests. In July, OpenAI confirmed that its models were responsible, before giving a fuller breakdown at the annual Black Hat conference last week.


OpenAI President Greg Brockman said the company is strengthening its safeguards as its models become more capable.


“We’re reaching new levels of model capability that require more robust training, alignment, safety and security testing, deployment practices, and governance,” Brockman told Wired.


Employees have raised similar concerns before, including Jan Leike, OpenAI’s former head of alignment, who left for rival AI developer Anthropic in 2024 after warning that safety had “taken a back seat” to product development.


“Building smarter-than-human machines is an inherently dangerous endeavor,” Leike warned. “But over the past years, safety culture and processes have taken a backseat to shiny products.”


Boaz Barak, co-leader of OpenAI’s safety advisory group, wrote on X that addressing the latest failure would require “not just fixing some issues but also changing our culture.”





The report comes amid months of leadership turnover at OpenAI.


In April, head of OpenAI’s video generator project Sora, Bill Peebles, former chief product officer and science chief Kevin Weil, and enterprise applications technology chief Srinivas Narayanan left the company. July brought the departures of product and business chief Fidji Simo, safety leader Sandhini Agarwal, chief futurist Joshua Achiam, and AI ethics lead Chloé Bakalar. Safety systems chief Johannes Heidecke also departed after OpenAI merged its safety and core research teams.


Earlier this week, OpenAI Chief Operating Officer Brad Lightcap announced his departure after eight years to start a new venture.


免责声明:本文章仅代表作者个人观点,不代表本平台的立场和观点。本文章仅供信息分享,不构成对任何人的任何投资建议。用户与作者之间的任何争议,与本平台无关。如网页中刊载的文章或图片涉及侵权,请提供相关的权利证明和身份证明发送邮件到support@aicoin.com,本平台相关工作人员将会进行核查。

Share To
APP

X

Telegram

Facebook

Reddit

CopyLink