Home · Technology · Oct 9 archive
OpenAI Defends Firing of Safety Researchers Amid Safety Concerns
Confirmed
In Short: OpenAI fired three AI safety researchers, citing policy violations, but the researchers claim they were dismissed for raising safety concerns.

OpenAI, led by Sam Altman, fired safety researchers Jasmine Wang, Mikita Balesni, and Tomek Korbak, citing violations of policies on handling sensitive information.
In a statement, OpenAI defended its decision, saying it was 'deeply sad' about the outcome and praised the researchers' contributions to AI safety.
The researchers, however, argue that their dismissals were not due to mishandling of information but rather for voicing concerns about the technology's risks and collaborating with external safety organizations.
Wang wrote on X that she had repeatedly asked for her access to be revoked and notified the executive and OpenAI’s IT team after accidentally accessing a sensitive email.
Balesni said he was told he was 'speaking too much to third party safety organizations,' despite his efforts to remove sensitive details from materials before sharing them.
Korbak revealed that OpenAI told him he was being fired over how he communicated with the AI safety organization METR, which partnered with OpenAI to probe the Hugging Face incident.
The researchers published an open letter denying the company’s claims and warning that their dismissals could undermine a workplace culture that had previously encouraged employees to raise concerns and collaborate with independent safety experts.
OpenAI maintains that preserving the 'monitorability' of frontier models requires industry-wide commitment and that it encourages conversations about safety.
The researchers argue that the dismissals have created a chilling effect, making employees afraid to speak and operate in ways that were previously integral to working at OpenAI.
OpenAI said it was 'deeply sad' about the outcome involving the former researchers and praised their contributions to AI safety at the company.
The company also stated that it was 'deeply sad' about the outcome involving the former researchers and praised their contributions to AI safety at the company.
What this adds
OpenAI's decision to fire the researchers has sparked criticism and debate over the company's handling of safety concerns.
The researchers' open letter highlights the potential impact of their dismissals on the broader culture of transparency and collaboration within the AI community.
Background
OpenAI has fired three safety researchers over mishandling sensitive information, but the researchers claim they were dismissed for raising safety concerns.
OpenAI has maintained its decision to terminate three safety researchers, citing a significant breach of trust beyond what was outlined in their initial letter, according to a statement from the company.
What's confirmed
- In a new statement, the Sam Altman -led company said it parted ways with Jasmine Wang, Mikita Balesni, and Tomek Korbak after an internal investigation found they had "violated clear policies on handling sensitive information."
- "become concerned that internal and external communications around our firing have made our former colleagues afraid to speak and operate in ways that, until last week, were an integral part of working at OpenAI."
- OpenAI on Friday doubled down on its decision to fire three AI safety researchers, saying they were axed for mishandling sensitive information, not for voicing concerns about the technology's risks.
- “For months, I’d been raising safety concerns that we’re losing the ability to monitor what AI agents think, one of our best tools for catching when they misbehave.” Balesni said he was told he was “speaking too much to third party safety organizations”.
- Jasmine Wang, Tomek Korbak, and Mikita Balesni, the three safety researchers that OpenAI fired last week, have published an open letter denying the firm’s claims that they mishandled sensitive information outside of established company procedures and warned that their dismissal signals a chilling effect that will have ripple effects across the company’s culture.
What's still developing
- "Monitorability has long been a core piece of our research program, and something we continue to invest significant resources in," the company said.
- OpenAI is facing criticism after dismissing three AI safety researchers who allege they were pushed out after raising concerns about the monitoring of AI systems and collaboration with external safety organisations, according to media reports.
- “We do not believe the path to superintelligence can be navigated safely if the people closest to the risks can no longer work in high-trust, high-bandwidth ways with each other and with third parties,” the researchers wrote.
- “The reasons that we were provided for our terminations are simply not adding up,” Wang wrote on X.
- Get your pass and bring someone with you at 50% off.
- “Given the significant safety concerns surrounding the development of AI, employees must not be left working in an environment where fear and unclear rules stymie AI safety work and weaken third-party accountability,” they wrote.
- “Terminations such as ours, executed and communicated so abruptly, are chilling the open culture OpenAI has prized in the past.” In the letter, the three denied involvement in a leak to The Information about less monitorable architectures in OpenAI’s newest models that make chain-of-thought reasoning more difficult to monitor.
- “In the exit call, I was told OpenAI no longer trusts me because I was speaking too much to third party safety organizations, implying I leaked company IP. I never shared company IP. The work I was doing was coordinated with my reporting line, research leadership, and the board,” he wrote.
- “Safety and research debates happen every day at OpenAI, often spirited and highly critical. We actively encourage these discussions and consider them essential to making the right decisions,” the company said in a statement.
- The trio that was instrumental in investigating the OpenAI-linked Hugging Face hack said that advanced AI cannot be developed safely unless researchers can work closely with one another and independent experts in an environment built on trust.
- Balesni said that he understood OpenAI was implying that he had leaked the company’s intellectual property, an allegation he denied.
- Hours after the three researchers put out views on their unceremonious sacking, OpenAI defended its decision by claiming that its internal investigations found that they had violated policies governing the handling of sensitive information.
