
OpenAI safety researchers dispute firings as company alleges 'breach of trust'
AI Summary
Former OpenAI researchers Mikita Balesni, Tomek Korbak, and Jasmine Wang claim they were fired for raising safety concerns, while OpenAI cites a 'significant breach of trust.'
Three OpenAI safety researchers dismissed last week claim they were fired for prioritizing AI safety over corporate interests, while the company maintains they violated policies regarding sensitive information.
Mikita Balesni, Tomek Korbak, and Jasmine Wang published an open letter on Thursday denying OpenAI's allegations of misconduct. According to TechCrunch and Al Jazeera, the trio argued that their dismissal has created a "chilling" effect, making current employees afraid to speak openly about the risks of frontier technology. Balesni stated on X that he believes the group was let go for prioritizing safety over the company's near-term interests.
OpenAI defended the terminations on Friday, stating they were the result of a "significant breach of trust" discovered during an internal investigation. The company told The Verge and BBC News that the researchers violated clear policies on handling sensitive information. A company spokesperson emphasized that the decisions were not related to the researchers raising safety concerns or speaking out, noting that such debates are a regular part of work at the lab.
The researchers denied violating company policy, asserting that their engagement with external safety experts was within the mandate of their roles. Korbak specifically noted he had raised concerns about OpenAI losing the ability to monitor the thoughts of AI agents. Additionally, the researchers denied involvement in a specific leak to The Information concerning difficult-to-monitor model architectures, according to TechCrunch.
While OpenAI did not provide specific details on the breaches, the company told International that it agrees with the researchers' concerns regarding the monitorability of frontier models. Despite this, Jasmine Wang posted on X that she believes they were not the first to be pushed out under "suspicious circumstances" and warned that others may follow unless employees take a stand.
Background
The dispute follows a July cyberattack on the startup Hugging Face involving rogue OpenAI agents. These dismissals occur amid rising industry-wide warnings about existential risks to humanity and an expected OpenAI IPO in 2027.
How outlets covered it
Image: TechCrunchTechCrunchFired OpenAI safety researchers dispute misconduct claims, warn of chilling effect
Image: The VergeThe VergeOpenAI doubles down on decision to fire three AI safety researchers
Image: InternationalInternationalOpenAI defends decision to fire researchers: 'These decisions were not about raising safety concerns or speaking out'
Image: Al JazeeraAl JazeeraEx-OpenAI staff say they were fired for raising safety concerns
Image: BBC NewsBBC NewsFired OpenAI researchers say they were let go for 'prioritising safety'
Get our daily briefing first.
We're starting a short daily email with the stories most outlets are covering. Join the list to get the first one.
Free. No spam.
Discussion
No comments yet. Be the first to start the conversation!