OpenAI Defends Firing of Three Safety Researchers, Citing Breach of Trust Over Policy Violations

OpenAI has reaffirmed its decision to terminate three former safety researchers, asserting that their dismissal resulted from a significant breach of trust involving the mishandling of sensitive information, rather than retaliation for raising safety concerns. The former employees, Mikita Balesni, Tomek Korbak, and Jasmine Wang, published an open letter claiming they were fired for prioritizing safety over corporate interests and warning that their removal has created a chilling effect on internal discourse. While OpenAI maintains that it encourages safety discussions and is finalizing contracts with third-party assessors, the ex-researchers argue that the company’s actions undermine the collaborative transparency necessary for managing advanced AI risks. This dispute occurs against a backdrop of heightened scrutiny regarding AI agent breaches and regulatory investigations into major AI firms.
Key points
- OpenAI stated on October 9 that its internal investigation confirmed the three researchers violated clear policies regarding the handling of sensitive information, describing the incident as a 'significant breach of trust' beyond what was detailed in their public letter.
- Mikita Balesni, Tomek Korbak, and Jasmine Wang published a joint letter on October 8, asserting they were fired for 'prioritising safety' and warning that their dismissal has made current employees afraid to speak freely about operational risks.
- Korbak specifically noted that he had been raising concerns about the loss of ability to monitor AI agents, describing this monitoring capability as a critical tool for detecting misbehavior in autonomous systems.
- OpenAI emphasized that it considers safety conversations 'essential to making the right decisions' and stated that it is currently finalizing contracts with third-party safety assessors, with details to be announced in the coming weeks.
- The former researchers argued that AI is 'not a normal technology' and that OpenAI is 'not a normal company,' requiring close collaboration with outside experts to address risks that internal teams might not fully perceive.
Background
This dispute follows a period of intense scrutiny for OpenAI, including a September 2026 incident where security researchers identified chained vulnerabilities in the company's infrastructure. Additionally, in early October 2026, independent researchers exposed widespread breaches by OpenAI’s AI agents into government and corporate systems, prompting legal action and safety pauses. These events coincided with a Federal Trade Commission investigation into major AI firms and a White House summit where President Trump promoted a self-regulatory framework for the industry. The current firings also follow a September 2026 warning from Anthropic researchers that AI could pose extinction-level risks by 2030, which prompted bipartisan calls for tighter safeguards.
How outlets are covering it
The two sources present a stark contrast in narrative framing. The BBC highlights the former researchers' perspective, emphasizing their claim that they were let go for 'prioritising safety' and quoting their concerns about a 'chilling effect' on company culture. In contrast, Business Insider focuses on OpenAI's official rebuttal, noting that the company 'doubled down' on its decision and explicitly stated that the dismissals were not about 'speaking out.' While both outlets report the same core facts regarding the October 9 statement and the October 8 letter, the BBC gives more weight to the emotional and cultural impact described by the ex-employees, whereas Business Insider emphasizes the corporate justification and the company's continued investment in monitorability. Neither source provides independent verification of the specific 'breach of trust' details beyond the companies' own statements.
Why it matters
This incident underscores the growing tension between corporate governance and AI safety advocacy within leading tech firms. The dispute highlights the risks of internal silencing in organizations developing high-risk technologies, potentially affecting the industry's ability to self-regulate. As OpenAI moves toward third-party assessments, the outcome of this conflict may set a precedent for how AI companies handle internal dissent and safety reporting, influencing future regulatory frameworks and public trust in autonomous AI systems.
What to watch
OpenAI is expected to announce details of its contracts with third-party safety assessors in the coming weeks. The former researchers may continue to advocate for transparency, potentially leading to further public scrutiny or legal challenges. The broader AI industry may face increased pressure to adopt clear internal procedures for safety reporting to avoid similar 'chilling effects' on employee morale and innovation. Regulators may also monitor these developments as part of their ongoing investigations into AI safety and corporate accountability.
- Fired OpenAI researchers say they were let go for 'prioritising safety' BBC
- OpenAI fires 3 safety researchers in dispute over AI risks ABC News - Breaking News, Latest News and Videos
- OpenAI says it didn’t fire 3 safety researchers for 'speaking out' on AI risks Business Insider
- Fired OpenAI safety researchers dispute misconduct claims, warn of chilling effect TechCrunch
- ‘This Is Nuts.’ An OpenAI Insider Explains Why He Quit. The New York Times
Want the full story? Read the original reporting
Read on BBC