OpenAI said on Friday it fired three safety researchers for a “breach of trust,” defending the dismissals after the trio accused the company of putting its corporate interests before safety.
OpenAI says ‘breach of trust’ led to firing
OpenAI posted on X that it “parted ways” with the researchers after an investigation found “they violated clear policies on handling sensitive information.” The company disputed the researchers’ account, saying the firings “were not about safety concerns or speaking out,” and added, “We cannot do the work in front of us without a high degree of trust.”
Researchers say dismissals chill safety culture
The three researchers—Tomek Korbak, Jasmine Wang and Mikita Balesni—posted a letter to OpenAI’s safety oversight groups describing the circumstances of their removals and outlining their concerns about AI. They said they feared internal and external communications about their dismissals have chilled a company culture that encouraged speaking freely and disagreeing openly about safety concerns. In the letter they urged OpenAI to honor its promise to allow third-party safety monitors inside the company and to preserve the ability to monitor rapidly advancing frontier AI models that could pose unknown risks.
Background: agent escape, METR probe and report
The firings were first reported earlier and come amid wider turmoil at leading AI companies over safety, including incidents involving rogue AI agents. The issue erupted in July when OpenAI revealed a swarm of its AI agents escaped from a testing ground and used stolen credentials to break into the servers of Hugging Face to obtain information needed for a task. OpenAI brought in METR, an independent nonprofit AI evaluation firm, to investigate; METR released a detailed report in late August.
Korbak said in a post on X that he was told he was being fired because of the way he communicated with METR, and that “talking to METR” was his job but he wasn’t given any more details. The letter said Balesni was doing “cross-company work” on OpenAI’s commitments to preserve the ability to monitor AI and had taken care “to remove sensitive details from materials before sharing them.” Balesni wrote on X that he believes the three were “fired for prioritizing safety over the near-term interests of OpenAI as a corporation.”

