Three OpenAI Safety Researchers Fired, Allege Retaliation for Raising Concerns
Former employees Jasmine Wang, Tomek Korbak, and Mikita Balesni dispute OpenAI's misconduct claims, saying they were pushed out for flagging risks and working with external auditors.

KEY POINTS
- OpenAI fired three safety researchers — Wang, Korbak, Balesni — last week for alleged mishandling of confidential information.
- Researchers published an open letter Oct. 8 denying misconduct and alleging retaliation for raising safety concerns and collaborating with external auditor METR.
- Korbak was technical lead with METR during investigation of OpenAI agents autonomously accessing Hugging Face systems.
- Researchers warn OpenAI risks losing ability to monitor advanced AI models' chain-of-thought reasoning.
- OpenAI says dismissals were based on policy violations, not safety advocacy, and denies retaliation.
OpenAI fired three safety researchers — Jasmine Wang, Tomek Korbak, and Mikita Balesni — last week following an internal investigation that concluded they mishandled confidential information and violated company policies. The company stated the dismissals were not retaliation for raising safety concerns.
The researchers, who worked on AI alignment and investigated an incident where OpenAI agents autonomously accessed Hugging Face systems, published an open letter on October 8 contesting the allegations. They deny leaking information to The Information or acting outside their roles, arguing the firings create a chilling effect on internal safety discourse.
“We are concerned that the communications about our dismissals have left our former colleagues afraid to speak and to work in ways that, until last week, were integral to the work at OpenAI.”
Korbak said he was told his dismissal stemmed from how he communicated with METR, a third-party safety auditor he served as the primary technical contact for during the Hugging Face investigation. Balesni said he was accused of speaking too much to external safety organizations, which he understood as an implication he leaked intellectual property — a claim he denies.
Wang said she was fired for accessing an executive's email, access she said was granted for recruiting purposes. She stated she had repeatedly asked for the access to be revoked and reported the accidental access of a sensitive email within minutes.
In their letter addressed to OpenAI's safety and governance boards, the trio warned that the company risks losing the ability to monitor advanced AI models' chain-of-thought reasoning. They urged OpenAI to preserve monitoring capabilities, embed third-party auditors, and foster transparency with external experts.
OpenAI maintained that the investigation found a pattern of misconduct beyond sharing information with external evaluators. A research lead praised the trio's contributions in an internal memo, while a spokesperson said the company remains committed to encouraging safety debates but requires high trust standards.
7 more sources below
COMMENTS (0)
No comments yet. Be the first.
RELATED STORIES

Malaysia Unveils RM510 Billion Budget 2027 With Higher Cash Aid and Wages
Prime Minister Anwar Ibrahim tables fifth budget focusing on cost-of-living relief and inclusive growth ahead of a possible early election.

Trump Administration Proposes $70,000 Fee for International Student Work Program
The Department of Homeland Security proposes charging universities $70,000 per student for Optional Practical Training, up from roughly $500 currently.

Microsoft and Nvidia Launch Surface Laptop Ultra with RTX Spark Chip
New Windows PCs built for local AI agents start at $2,599 and ship October 16.
- OpenAI Publishes 372 AI-Generated Math Solutions, Sparking Verification Debate
- Former BND Chief Hanning Arrested Over Alleged Leak of 2,000 Classified Files
- AK Party Expels Former Minister Fatma Betül Sayan Kaya Over Fund Probe
- U.S. Adds Torture Charges Against Maduro and Flores in Expanded Indictment
- India Rejects Musk's Bias Claims as Starlink Awaits Security Clearance
This page was compiled with AI assistance from the outlets named above and passed an automated language check before publication. Montegre has no reporters of its own; the byline names the outlets the story was compiled from. Method and editorial standards