Tech
EN AZ
Fired OpenAI safety researchers dispute misconduct claims, warn of chilling effect

Fired OpenAI safety researchers dispute misconduct claims, warn of chilling effect

techcrunch.com 08.10.2026 22:04 7 views
Three fired OpenAI safety researchers dispute allegations of mishandling sensitive information, warning in an open letter that their dismissals are creating a chilling effect on the company’s AI safety culture.

Jasmine Wang, Tomek Korbak, and Mikita Balesni, the three safety researchers that OpenAI fired last week, have published an open letter denying the firm’s claims that they mishandled sensitive information outside of established company procedures and warned that their dismissal signals a chilling effect that will have ripple effects across the company’s culture. The researchers were dismissed last week after allegedly sharing confidential company information with a third-party AI safety organization. OpenAI said at the time that they violated the company’s policies by “accessing and handling sensitive company information.” “AI is not a normal technology, and OpenAI is not a normal company,” Wang, Korbak, and Balesni wrote.

The freedom to do so without fear, and to have well-defined internal procedures that enable this work, is itself an essential safety mechanism.” They said that their firing represents a broader shift in the culture of OpenAI, one that used to encourage workers to “raise safety concerns and disagree openly.” They said employees are now “unclear on where they stand” when behavior that was allegedly normal a month ago is now suddenly grounds for dismissal. They also denied engaging with external parties outside the mandates of their jobs. OpenAI has not formally responded to the open letter, but shared with TechCrunch an internal memo attributed to a research leader, praising the three researchers’ contributions to AI safety and denying that they were fired in retaliation.

We do not terminate employees for raising concerns.” OpenAI did not directly address TechCrunch’s questions about which policies the researchers allegedly violated, the circumstances of their dismissal, or how the company protects employees who raise safety concerns and collaborate with external evaluators. The firings have fueled speculation about their circumstances, particularly as OpenAI faces scrutiny over recent safety incidents involving rogue agents and leaks about its models. The letter also addresses the researchers’ response to the Hugging Face incident, in which a swarm of agents broke out of their sandbox and breached external systems.

The letter says that the incident and investigation was “without precedent,” meaning “internal policies were being developed in real time.” Due to the sensitive nature of the investigation, Korbak believed he was acting within OpenAI’s policies and norms by communicating closely with outside safety evaluators to build trust, per the letter. At the same time, Balesni was also working internally to address the growing AI monitorability problem, an effort the researchers say in their letter “can only succeed through extensive communication with external parties.” According to the letter, Balesni coordinated with and was supported by OpenAI board members and executives throughout his work. They did not action my request, I couldn’t remove it myself, and the inbox was combined in an indistinguishable way in my phone’s mail app.

When I opened a sensitive email by mistake, I told the executive within minutes and asked IT again. None of this was hidden.” Wang went on to say that the reasons behind the terminations are “not adding up,” and that she and her colleagues are “not the first to be pushed out of OpenAI under suspicious circumstances.” The researchers called on OpenAI to adhere to its public commitments to embed third-party safety auditors within the organization, to preserve monitorability of frontier models, and “continue to support an open and transparent culture of dialogue between safety researchers and the rest of the safety ecosystem.” OpenAI agrees with their recommendations, per the memo. You can’t build AGI safely if the people closest to the risks are afraid to speak.”

Extract — continue reading at the source.

Read full story