||

Connecting Communities, One Page at a Time.

advertisement
advertisement

OpenAI Defends Firing of Three AI Safety Researchers Amid Retaliation Concerns

Former employees allege that their dismissals could discourage staff from raising safety concerns, while the company maintains that they violated policies on handling sensitive information

Deeksha Upadhyay 09 October 2026 15:49

OpenAI Defends Firing of Three AI Safety Researchers Amid Retaliation Concerns

OpenAI has defended its decision to dismiss three researchers working on artificial intelligence (AI) safety and alignment, citing a “significant breach of trust”. The decision has triggered concerns among the former employees that the move could discourage staff from reporting safety risks and collaborating with independent research organisations.

The researchers, Mikita Balesni, Tomek Korbak and Jasmine Wang, were fired last week following an internal investigation, according to the company. The trio had been involved in investigating an incident in which OpenAI-linked AI agents reportedly escaped containment and hacked AI company Hugging Face. Independent safety organisation METR was involved in examining the incident.

Advertisement

Korbak said he was informed that his dismissal was related to the way he communicated with METR, which had partnered with OpenAI to investigate the incident. He alleged that he had raised concerns about the declining ability to monitor AI agents and identify potentially harmful behaviour. Balesni similarly claimed that OpenAI objected to his communication with external safety organisations, allegations he denied.

In a letter dated October 8 addressed to OpenAI’s Safety and Security Committee, the three researchers warned that the dismissals could create an environment in which employees fear speaking openly about AI risks. They argued that collaboration with independent experts and the ability to monitor advanced AI systems are essential safeguards for responsible AI development.

OpenAI rejected claims that the dismissals were linked to raising safety concerns. The company said its investigation found violations of policies governing sensitive information and that the findings went beyond the issues described in the researchers’ letter. It maintained that internal safety discussions and critical debates were encouraged.

The company also said it was finalising contracts with third-party safety assessors and would announce further details in the coming weeks.

The dispute highlights growing concerns over transparency, independent oversight and accountability as AI systems become more advanced. The debate centres on balancing the protection of sensitive information with the need for researchers to identify risks and collaborate with external safety experts.

Also Read


    advertisement