Three Different People, Three Different Stories, One Disputed Verdict
OpenAI's public explanation for the firings is deliberately generic: a 'thorough investigation found they violated clear policies on handling sensitive information' [1]. But each of the three researchers was reportedly given a different, highly specific personal explanation, and none of the three maps cleanly onto the other two. Korbak says he was told, verbally, that he was fired 'because of the way I communicated with METR' [2]- the external auditor examining the Hugging Face containment breach. Wang says her dismissal was tied to access to an executive's email that had been delegated to her for recruiting purposes and that IT never revoked [2]. Balesni, for his part, denies being the source of a separate leak to The Information about less-monitorable model architectures, saying he removed sensitive details before sharing anything and acted within company norms at the time [3].
That divergence matters because it undercuts the idea of a single, coherent policy violation. If three people were fired for the 'same' breach of trust, you would expect one shared factual story, not three unrelated ones that each researcher individually contests. OpenAI's rebuttal does not resolve this; instead it adds a fourth, unfalsifiable layer, stating its internal investigation uncovered 'a significant breach of trust beyond what's outlined in the letter they published' [2]- a claim that cannot be checked against anything public. The result is a dispute where the only verifiable common thread connecting all three firings is timing, not substance.


