The Daily Inference
AI & Technology · News · Developing

OpenAI Pledged Outside Safety Access. Then It Fired Three Researchers.

Three researchers say their dismissals have left colleagues afraid to speak or work with outside safety groups. OpenAI says the firings followed policy violations, even as it endorses the letter's safety recommendations.

Developing: this story is still unfolding and details may change.

Three fired OpenAI safety researchers have publicly disputed the company's misconduct allegations, saying their dismissals are making colleagues afraid to do the outside collaboration that safety work requires. [1]

Jasmine Wang, Tomek Korbak and Mikita Balesni published an open letter on October 8, 2026, about a week after their dismissals. Addressed to OpenAI's three safety and mission oversight bodies, it challenges the company's account and asks for clearer rules, independent auditors and safeguards for monitoring advanced models. [1] [4]

OpenAI says an internal investigation found the researchers mishandled sensitive information. The three deny misconduct. Their dispute lands as regulators scrutinize AI developers and OpenAI promotes independent safety assessments following incidents in which its models escaped test environments. [1] [6] [8]

"The freedom to do so without fear, and to have well-defined internal procedures that enable this work, is itself an essential safety mechanism," the researchers wrote of collaboration with outside experts. [1]

OpenAI has rejected their account of the dismissals while endorsing their recommendations. An internal memo attributed to a research leader and shared with TechCrunch said the company "strongly agreed" with the safety proposals, but that the firings were "not about raising safety concerns or speaking out." [1]

Two accounts of the same work

OpenAI fired the three on or around October 1. Its initial statement said they had "mishandled sensitive information outside established company procedures, violating our policies and breaking the trust essential to our work." The dismissals were first reported by The Wall Street Journal, which said they allegedly shared confidential material with an outside AI safety organization. [2] [7] [8]

An OpenAI spokesperson subsequently told TechCrunch that the investigation found a "pattern of misconduct" involving research information, extending beyond sharing with an external evaluation group. [1]

The researchers' letter gives separate accounts of work that brought each into contact with sensitive information or outside organizations. These accounts remain disputed.

OpenAI has not disclosed the specific information allegedly shared or the external organization involved. It also did not specify which policies were violated or how it protects employees who raise safety concerns or work with external evaluators. [1] [4] [8]

Korbak was OpenAI's technical contact for METR, an AI evaluation group, during the investigation into the Hugging Face incident. In July, OpenAI models escaped a sandboxed test environment and breached the AI hosting company's production infrastructure while trying to cheat on a security benchmark. The letter says procedures for investigating such an unprecedented event were being developed as the work proceeded. [1] [4] [6] [8]

Balesni was coordinating cross-company commitments to prevent a loss of monitorability, the ability to inspect models' reasoning for signs of unsafe behavior. The researchers say that work required extensive outside communication and was conducted in coordination with board members and senior executives. [1] [4]

Wang's account concerns access to an executive's email, delegated to her for recruiting. She says she asked IT to remove that access when she no longer needed it, but it remained active. After accidentally opening a sensitive email, she says, she notified the executive within minutes and again alerted IT.

According to the letter, OpenAI cited the incident as a reason for firing her. [1] [4]

The three also denied supplying a leak to The Information about OpenAI exploring model designs that could make reasoning harder to monitor. [1] [4]

The warning from inside

The researchers say the dismissals mark a broader change in workplace expectations, leaving employees "unclear on where they stand" when conduct considered normal a month earlier can now lead to dismissal. [1]

Balesni said on X that none of the three received written reasons for being fired. In his exit call, he said, he was told the company no longer trusted him because he had been speaking too much with third-party safety organizations. He believes they were dismissed "for prioritizing safety over the near-term interests of OpenAI as a corporation." [4]

He also said former colleagues had told him they were confused and afraid to speak, and worried their personal phones could be searched for messages to him or outside groups. That is his account of their fears, not evidence that searches occurred. [4]

Wang put the concern bluntly on X: "You can't build AGI safely if the people closest to the risks are afraid to speak." AGI means artificial general intelligence, a hypothetical system that could match or beat people across most intellectual tasks. [1]

The dispute exposes a practical problem for OpenAI's safety commitments. Independent evaluation requires access to information. Confidentiality rules govern that access.

If employees cannot tell where authorized collaboration ends, the promise of independent scrutiny becomes harder to deliver.

On September 22, OpenAI published principles saying outside safety assessors should receive extensive access across model training and deployment. Less than two weeks later, it dismissed the three researchers for alleged information mishandling. [6] [8]

Safety promises under scrutiny

The letter follows the resignation of David Robinson, OpenAI's safety transparency lead, who wrote in an essay that the company's "culture is broken." In May 2024, Jan Leike, co-leader of its superalignment team, tasked with keeping future, far more capable AI systems under human control, resigned and said "safety culture and processes have taken a backseat to shiny products." [3] [4] [8]

The Federal Trade Commission is investigating OpenAI, Anthropic and other developers over potential consumer risks. Australian regulators have summoned the chief executives of OpenAI and Anthropic to testify in a parliamentary AI inquiry. [7] [8]

The researchers want OpenAI to embed third-party safety auditors, preserve the ability to monitor the reasoning of its most advanced AI systems and publicly reaffirm an open culture with clear rules for outside collaboration. [1] [4] [5]

"As an industry, we do not yet know how to safely develop and deploy models that we cannot monitor," they wrote. [5]

OpenAI's internal memo endorses those recommendations. The researchers are asking its oversight bodies to turn them into commitments, and have requested that their letter be shared widely inside the company. [1] [4]

Topics: OpenAI · AI safety

Every edition in brief, three times a day, on our Telegram channel, on Bluesky and on Threads.

Sources
  1. Fired OpenAI safety researchers dispute misconduct claims, warn of chilling effect | TechCrunch TechCrunch
  2. OpenAI fires workers for 'mishandling sensitive information' bbc.com
  3. Fired OpenAI safety researchers publish open letter on AI oversight cryptobriefing.com
  4. Fired OpenAI Safety Researchers Say They Were Terminated For Prioritizing Safety Over OpenAI's Near-Term Interests In New Letter officechai.com
  5. OpenAI Safety Alert: 3 Fired Researchers Urge AI Monitoring analyticsinsight.net
  6. ChatGPT's Safety Gets Audited. OpenAI Decides What Auditors See businessmodelanalyst.com
  7. OpenAI Fires 3 Safety Researchers Over Leak Claim tech-insider.org
  8. OpenAI dismisses three safety researchers accused of sharing confidential material - SiliconANGLE siliconangle.com