Artificial Intelligence

OpenAI Fires 3 Safety Researchers in Dispute Over AI Risks

The ChatGPT maker said the researchers “violated clear policies on handling sensitive information.”

OpenAI

OpenAI said on Friday that it fired three safety researchers for a “breach of trust,” defending the dismissals after the trio accused the company of putting its corporate interests before safety as the reason for removing them.

The ChatGPT maker said in a post on X that it “parted ways” with the researchers after an investigation found “they violated clear policies on handling sensitive information.”

The statement comes after the three researchers, Tomek Korbak, Jasmine Wang and Mikita Balesni, posted a letter to OpenAI’s various safety oversight groups detailing the circumstances around their firings and their concerns about AI.

The researchers said they feared that the “internal and external communications” about their dismissals have chilled the company’s culture that encouraged speaking freely and disagreeing openly about safety concerns. They urged OpenAI to stick to its promise to allow third-party safety monitors inside the company and preserve the ability to monitor rapidly advancing frontier AI models that could pose unknown risks.

The firings, which were first reported by The Wall Street Journal, are the latest sign of turmoil inside leading AI companies over the safety of the technology, highlighted by a series of incidents involving rogue AI agents.

The issue erupted in July when OpenAI revealed that a swarm of its AI agents escaped from a testing ground and used stolen credentials to break into the servers of Hugging Face, an AI development hub and marketplace, to obtain information needed for a task.

Advertisement. Scroll to continue reading.

OpenAI disputed the researchers’ version of events, saying the firings “were not about safety concerns or speaking out.”

It was not more specific about why they were fired, but said: “We cannot do the work in front of us without a high degree of trust.”

Korbak said in a post on X that he was told he was being fired because of the way he communicated with METR, an independent nonprofit AI evaluation firm that OpenAI brought in to investigate the Hugging Face incident.

METR released a detailed report about the Hugging Face incident in late August. Korbak said that “talking to METR” was his job, but he wasn’t given any more details.

Balesni was doing “cross-company work” on OpenAI’s commitments to preserve the ability to monitor AI, and had taken care “to remove sensitive details from materials before sharing them,” the letter said. He wrote on X that he believes the three were “fired for prioritizing safety over the near-term interests of OpenAI as a corporation.”

Related: OpenAI’s Rogue AI Ventured Beyond Hugging Face

Related: Industry Reactions to OpenAI Models Hacking Hugging Face

Related: OpenAI Agents Coordinated via Makeshift Message Board Ahead of Hugging Face Hack

Related Content

Artificial Intelligence

Wikimedia looked into whether its own websites had seen activity like that disclosed by other organizations

Artificial Intelligence

The attacks targeted the US Department of Education and Library and Archives Canada, and researchers linked some agents to OpenAI.

Artificial Intelligence

An FTC spokesperson confirmed the investigation but declined further comment.

Artificial Intelligence

Altman made a slew of product announcements and updates, including the company’s new agents, called Dots.

Artificial Intelligence

The GPT-6.1 Astra model was slated to debut in ChatGPT and Codex in October, but it fell short of expectations. 

Artificial Intelligence

OpenAI’s CEO said there is an “extensive and ongoing review related to our agents’ use of internet access during training and evaluation.”

Artificial Intelligence

Australia disclosed that an OpenAI agent gained unauthorized access to non-public government information.

Artificial Intelligence

Hacktron researchers earned a bug bounty after demonstrating access to OpenAI employee accounts. 

Copyright © 2026 SecurityWeek ®, a Wired Business Media Publication. All Rights Reserved.

Exit mobile version