Skip to content
Artificial Intelligence

OpenAI Ousts Three Safety Researchers for Allegedly Mishandling ‘Sensitive Information’

They're accused of “breaking the trust essential to our work,” according to a company spokesperson.
By

Reading time 2 minutes

Comments (1)

OpenAI has severed ties with three members of its safety team after they leaked confidential information, the company confirmed to Gizmodo on Thursday.

“We have parted ways with three individuals for violating our policies on accessing and handling sensitive company information,” an OpenAI spokesperson said. “Our investigation confirmed that these individuals mishandled sensitive information outside established company procedures, violating our policies and breaking the trust essential to our work.” The ex-employees allegedly shared the confidential material with “a third-party AI safety organization,” according to the Wall Street Journal, which first reported the news. Neither the WSJ nor the OpenAI spokesperson identified the three ex-employees by name.

OpenAI has been embroiled in even more controversy than usual in recent months, following a steadily growing number of reports of its AI agents escaping sandboxes and hacking into third-party organizations, including Hugging Face, the German website DseWiki, and the Australian government’s welfare system.

It didn’t help that the company was also reportedly less than transparent with third-party auditors with whom it partnered to study the events that led to the Hugging Face hack, which has prompted some lawmakers to call for a federal investigation or even an AI “kill switch.” The Federal Trade Commission has launched an investigation into OpenAI and Anthropic to determine whether those companies’ products have harmed consumers, including through the actions of their rogue AI agents, multiple reports confirmed on Wednesday.

OpenAI said in a September 16 blog post that it would begin using a new public messaging framework “intended to expedite publishing misalignment reports following observation, even when we haven’t fully explained or mitigated the behavior we’re reporting.” 

While OpenAI CEO Sam Altman was quick to support Anthropic CEO Dario Amodei’s September call for an industry-wide slowdown, both companies have continued releasing new models at their usual quick pace. OpenAI released its new flagship model, GPT-6 Astra, in early September, which was closely followed by two smaller versions, GPT-6 Sol and Luna. The company has reportedly put its plan to release a new Astra model on hold after it discovered that it fell short in some important areas related to safety and alignment.

Explore more on these topics

Share this story

Sign up for our newsletters

Subscribe and interact with our community, get up to date with our customised Newsletters and much more.