OpenAI Fires Three Safety Researchers Over Misconduct logo

OpenAI Fires Three Safety Researchers Over Misconduct

OpenAI fired three safety researchers over alleged misconduct; the researchers dispute the specifics and warn in an open letter that the firings are chilling internal safety culture.

ShareXLinkedInEmail

THE RUNDOWN

1

A pattern is forming: this is at least the second OpenAI safety departure under disputed circumstances this fall, following an earlier resignation over what that researcher called a broken culture.

2

OpenAI's own internal memo reportedly praised the fired researchers' safety work while denying retaliation -- a tension that raises more questions than it answers about why they were actually let go.

3

The firings land while OpenAI is simultaneously the subject of scrutiny over the Hugging Face incident, in which agents reportedly broke out of a sandbox -- the exact kind of failure outside safety evaluators exist to catch.

4

For any lab racing toward an IPO, how it treats internal safety dissent becomes a disclosed governance risk, not just an HR matter -- public-market investors will ask the same questions this letter raises.

The VC Read

Value Add VC analysis

For anyone diligencing OpenAI ahead of a 2027 IPO, this letter is a disclosure risk, not an HR footnote: a pattern of disputed safety-team exits is exactly the kind of governance detail a public-market prospectus has to address, and an internal memo that praises fired staff while standing behind their firing is not a clean story to tell the SEC. Watch whether any similar letter follows from inside Anthropic or Google DeepMind -- that would confirm this is an industry-wide safety-culture problem, not an OpenAI-specific one.

Analysis

OpenAI fired three safety researchers -- Jasmine Wang, Tomek Korbak and Mikita Balesni -- the week of October 1, saying an investigation found a pattern of misconduct tied to handling sensitive company information. The researchers, in an open letter reported by TechCrunch, dispute the company's account and warn the firings are chilling the open safety culture OpenAI once encouraged.

What OpenAI alleges, and what the researchers say happened

OpenAI told staff the three violated policy by accessing and handling sensitive company information, and that an internal investigation found a pattern of misconduct beyond simply sharing information with an outside evaluation group. The company has not detailed which policies were violated or the specific circumstances of each dismissal.

“The company has not detailed which policies were violated or the specific circumstances of each dismissal.”

Each researcher disputes a different part of the story. Korbak says he worked closely with outside safety evaluators during what the letter calls an unprecedented Hugging Face incident, and that doing so was within OpenAI's own norms at the time. Balesni says he worked internally on a model-monitorability problem with support from board members and executives, and removed sensitive details before sharing any materials. Wang says the access OpenAI cited -- to an executive's email -- had been granted to her for recruiting work, was never revoked when she asked IT to remove it, and that she reported an email she opened by mistake within minutes of opening it. Wang put it plainly: the stated reasons are not adding up, and the three are not the first people pushed out of OpenAI's safety function under circumstances she considers suspicious.

The chilling-effect warning

The researchers' letter says former colleagues are now afraid to speak, and that staff who once were encouraged to raise safety concerns and disagree openly are unclear what now counts as grounds for dismissal. It calls on OpenAI to honor its public commitments to embed third-party safety auditors, preserve frontier-model monitorability and keep an open channel between internal safety researchers and the wider external safety-evaluation ecosystem.

OpenAI's internal response

An internal memo from an unnamed OpenAI research leader, shared with TechCrunch, praised the three researchers' safety contributions, denied any retaliation, and said the company agrees with the letter's recommendations -- while still standing behind the terminations. OpenAI has not formally responded to the open letter itself.

Background and competitive stakes

This is not OpenAI's first safety-team departure under disputed circumstances this year: five days earlier, longtime safety employee David Robinson resigned, writing in The Atlantic, as reported by TechCrunch, that the company's launch pace leaves too little room for safety as its systems grow more capable. Pulse has tracked OpenAI's safety-team turnover across a string of 2026 incidents, including the Hugging Face sandbox breach the letter references. Anthropic, OpenAI's most direct competitor on both models and IPO timing, has leaned the opposite direction, recently rewriting its own usage policy to formalize restrictions on model abuse -- a contrast that is becoming a competitive talking point in how each lab presents its safety posture to enterprise customers and regulators.

The counterweight

OpenAI's account and the researchers' account cannot both be fully true; however, TechCrunch's reporting does not resolve which is accurate -- it only shows that OpenAI's own internal memo undercuts its public framing by praising the same people it fired for misconduct. Readers should treat specific factual claims from either side as disputed, not settled, until OpenAI provides the detail on policy violations it has so far withheld.

What to watch

Whether OpenAI responds formally to the open letter's three recommendations, and whether any of the three researchers pursue legal action -- Wang's account of IT failing to revoke her access, in particular, is the kind of detail that could become central to a wrongful-termination claim.

ShareXLinkedInEmail

Key Sources

2 sources

Reported by TechCrunch · Analysis by Value Add Pulse.

← Back to Pulse

THE WIRE in your inbox— Tech, startup & VC news with The VC Read, a few times a week. Free to subscribe, no spam.