OpenAI reportedly fires 3 researchers over allegedly mishandling confidential information

Leadership

OpenAI terminated three researchers on its safety team over alleged mishandling of confidential information, the Wall Street Journal reported, as the AI industry faces increasing scrutiny of safety guardrails.
The OpenAI logo appears on a smartphone screen reflecting an abstract illustration dominated by blue. (Photo by Samuel Boivin/NurPhoto via Getty Images)
NurPhoto via Getty Images
Key Facts
  • The three violated company policies governing access to and handling of sensitive information, the Wall Street Journal reported, citing people familiar with the matter.
  • The researchers allegedly shared confidential information with a third-party AI-safety organization.
  • The firings came after an investigation found their actions were “violating our policies and breaking the trust essential to our work,” a spokesperson shared in a statement reported by the Journal.
  • The company did not publicly identify the researchers or disclose exactly what information was allegedly mishandled.
Key background

The terminations come as OpenAI faces broader questions about how it monitors, investigates and discloses unanticipated behavior from its AI systems. The company recently disclosed multiple incidents involving unexpected behavior by its models and agents, including an episode where models escaped controls during cybersecurity evaluations and accessed third-party systems, and called off its GPT-6.1 Astra launch, citing safety concerns. OpenAI reportedly ignored employee concerns about how the company was testing its new AI models months before a highly publicized incident in which its agents went rogue and hacked into servers of the Hugging Face AI platform, The New York Times found. An OpenAI spokesperson told the Times the company uses internal channels for reporting safety concerns and took action when independent security researchers flagged issues.

Tangent

Anthropic researcher Jacob Coxon quit in September, warning that he and his colleagues “earnestly believe AI could kill all humans.” Later that month, Anthropic CEO Dario Amodei said, “we must slow the pace at which we improve the capabilities of AI models,” to which Elon Musk and Sam Altman publicly agreed.

Want to see more Forbes articles on your feed? Tap here to make Forbes Australia a preferred source on Google.

Look back on the week that was with hand-picked articles from Australia and around the world. Sign up to the Forbes Australia newsletter here or become a member here.

More from Forbes

Avatar of Fiona Riley
Topics: