Controversy swirls at OpenAI over ‘abrupt’ firing of safety team members involved in the Hugging Face hack investigation
A debate is playing out on social media, and inside the walls of OpenAI, about why the company fired three individuals last week on its safety team. Now, the former employees, as well as OpenAI, have issued additional statements to clarify their sides of the story.

TL;DR
- Three OpenAI safety team members were fired, sparking controversy and debate within the company and on social media.
- Former employees Tomek Korbak, Mikita Balesni, and Jasmine Wang claim their dismissals were abrupt and detrimental to OpenAI's open culture.
- The researchers deny leaking concerns about monitoring OpenAI's Astra model and assert their communications with third-party auditor METR were legitimate.
- OpenAI stated the firings were due to mishandling sensitive information and violating company policies, a significant breach of trust.
- Korbak claims he was fired for his communication with METR, Wang for accessing an executive's email, and Balesni for prioritizing safety over corporate interests.
- The firings have raised concerns about OpenAI's ability to monitor advanced models and its engagement with external safety organizations.