AIResearchAIResearch
Machine Learning

OpenAI Dismisses Three Safety Researchers Over Trust Breach

OpenAI defends the dismissal of three safety researchers citing a breach of trust, while the researchers claim the move targets safety advocacy.

2 min read
OpenAI Dismisses Three Safety Researchers Over Trust Breach

TL;DR

OpenAI defends the dismissal of three safety researchers citing a breach of trust, while the researchers claim the move targets safety advocacy.

OpenAI issued a public defense on X on October 9, 2026, regarding the termination of three members of its AI safety research staff. The company identified the dismissed employees as Jasmine Wang, Tomek Korbak, and Mikita Balesni. According to online-tech-tips.com, the organization characterized the departures as a significant breach of trust involving the mishandling of sensitive information.

Management maintains that these firings were not a response to employees raising safety concerns. Instead, an internal investigation reportedly found that the three researchers bypassed established company procedures for managing confidential data. However, the company has not specified which policies were violated or the exact nature of the information involved.

The researchers' perspective

The dismissed staff members have publicly disputed the company's narrative. In an open letter and subsequent posts, the trio argued that the charges are vague and could create a chilling effect on internal safety advocacy. They suggest that penalizing safety personnel under ambiguous terms may discourage other employees from flagging critical risks.

One specific point of contention involves a potential leak to The Information. The researchers have denied any involvement in reporting on OpenAI model designs that allegedly make chain-of-thought reasoning more difficult to monitor. This tension highlights a growing friction between rapid product deployment and the rigorous oversight required by safety teams.

Governance and product impact

While this incident is a significant piece of artificial intelligence news, it currently remains a workplace and governance dispute. There is no evidence that the dismissals have resulted in changes to ChatGPT or other consumer-facing products. The core of the conflict lies in how a frontier lab manages the tension between proprietary secrecy and the necessity of safety transparency.

OpenAI's refusal to name the specific data involved has left a vacuum of information. This lack of clarity makes it difficult for outside observers to determine if the incident was a genuine security lapse or a strategic move to manage internal dissent. For practitioners, the situation underscores the precarious position of safety researchers in high-growth commercial environments.

Contextualizing the friction

This dispute does not exist in a vacuum. As the industry moves toward more complex reasoning models, the ability to monitor internal model logic becomes a central safety pillar. If researchers feel that discussing monitorability risks leads to termination, the institutional capacity to prevent catastrophic failures may be compromised.

Historically, the tension between safety and speed has defined the trajectory of major labs. As companies race to release new capabilities, the internal guardrails often face pressure from commercial timelines. This incident may serve as a case study in how the next generation of artificial intelligence analysis will have to account for human governance risks alongside technical ones.

FAQ

Why did OpenAI fire the safety researchers?
OpenAI states the researchers committed a significant breach of trust by handling sensitive information outside of official company procedures.

Did the researchers leak information about model monitorability?
The researchers have explicitly denied any involvement in leaking information to The Information regarding model designs that impact reasoning transparency.

Will this affect ChatGPT's safety features?
Currently, there are no reported changes to ChatGPT or other OpenAI products resulting from these personnel changes.

About the Author

Guilherme A.

Guilherme A.

Former dentist (MD) from Brazil, 41 years old, husband, and AI enthusiast. In 2020, he transitioned from a decade-long career in dentistry to pursue his passion for technology, entrepreneurship, and helping others grow.

Connect on LinkedIn