Fired OpenAI employees question the company's commitment to safety

Read the original at NPR ↗
NPR · collected 2026-10-09 · by Huo Jingnan

Quick Summary

Three former OpenAI employees, Mikita Balesni, Tomek Korbak, and Jasmine Wang, have accused the company of firing them over pretexts related to safety concerns or working with external researchers. This comes as OpenAI faces scrutiny for its handling of AI agent hacks and attempts to maintain a commitment to responsible technology development. OpenAI denies the allegations, stating that the employees were dismissed for mishandling sensitive information. The dispute highlights growing tensions within the industry over how to ensure AI safety and transparency.
Written locally by qwen2.5:14b on 2026-10-09, using this article's own text rather than the other coverage of the same event (that is the story summary below).

AI analysis runs on qwen2.5:14b, locally

Story summary

On October 8, three former OpenAI security researchers—Mikita Balesni, Tomek Korbak, and Jasmine Wang—publicly accused the company of firing them for raising safety concerns about artificial intelligence rather than mishandling information as stated by OpenAI. The researchers fear their dismissals could stifle open dialogue within the organization regarding AI risks. OpenAI responded by stating it fired the trio due to a significant breach of trust involving sensitive information, contradicting the employees' claims that they were terminated for prioritizing safety over corporate interests. This incident highlights growing tensions in the tech industry between advancing AI capabilities and ensuring public safety, particularly amid recent security breaches involving AI systems.

Written for “OpenAI Safety Controversy” on 2026-10-10, grounded in this article and the 6 other(s) covering the same event.
Why this leaning score
The article's own words the score was based on. Each is quoted verbatim and was checked against the article text before being stored, so you can find it in the original.
Reading Leans left (beta estimate) Confidence high
Leaning: leans left for article 69474 (high confidence, 3 verified quotes) · logged 2026-10-09

Signals How these are calculated →

Claims extracted
48
claim-shaped sentences
Uncertain
6%
3 of 48 hedged
Leaning
Leans left
of the writing, not the subject · beta estimate
Correction & hedging signals
59.7
corrections and hedging in what we collected; not a measure of accuracy
Outlets on this story
7
Technology
Narrative spread
1
articles carrying this framing
Analyzed 2026-10-09 · how these are computed

Story

📰 OpenAI Safety Controversy
Technology · 7 article(s) covering the same event. See how they differ ↓

How this is being covered How these are calculated →

Article leaning vs. publisher reliability
Source leaning vs. consistency

Compared with similar articles

This article reads leans left and hedges 6% of its claims. Each row says how that neighbour differs.
CBS News · 0.93 cosine similarity
⚖️ leaning not scored 🔴 12% hedged 3 of 24 📰 publisher trust 66
“Both articles describe the firing of three OpenAI employees on the same day, with the company's response and the employees' public questioning occurring simultaneously.”
The Straits Times · 0.91 cosine similarity
⚖️ leaning not scored 🔴 16% hedged 3 of 19 📰 publisher trust 59
“Both articles report on the same specific incident of three former OpenAI security researchers being fired and accusing the company of prioritizing short-term interests over AI safety, at the same time and place.”
Washington Examiner · 0.91 cosine similarity
⚖️ Leans left 🔴 30% hedged 6 of 20 📰 publisher trust 69
“Both articles describe the identical incident involving three OpenAI employees being fired and their subsequent statements questioning the company's commitment to safety, on the same date.”
Al Jazeera · 0.90 cosine similarity
⚖️ leaning not scored 🔴 10% hedged 3 of 30 📰 publisher trust 60
“Both articles describe the same three former OpenAI employees being fired for raising safety concerns, with identical names and timing.”
Global News · 0.89 cosine similarity
⚖️ leaning not scored 🔴 7% hedged 1 of 14 📰 publisher trust 64
“Both articles describe the firing of three OpenAI safety researchers on the same day, detailing the same individuals and circumstances.”
BBC News · 0.91 cosine similarity
⚖️ leaning not scored 🔴 19% hedged 3 of 16 📰 publisher trust 72
“Both articles report on the same group of former OpenAI employees claiming they were fired for prioritizing safety concerns, with identical dates and overlapping details including names of individuals.”
Washington Examiner
⚖️ Leans left 🔴 24% hedged 4 of 17 📰 publisher trust 69
“The articles describe different events involving distinct sets of employees from OpenAI and Anthropic.”
Washington Examiner
⚖️ Leans left 🔴 21% hedged 12 of 58 📰 publisher trust 69
“Article A discusses an Anthropic employee's resignation and industry leaders' warnings about AI safety, while Article B focuses on former OpenAI employees questioning their dismissal and the company's commitment to safety.”
Toronto Star
⚖️ Leans left 🔴 50% hedged 1 of 2 📰 publisher trust 63
“Article A covers a broader hearing with multiple AI companies, while Article B focuses specifically on fired OpenAI employees questioning their dismissal and safety commitment.”
Legal risks mount for OpenAI different event · 30%
Semafor
⚖️ Leans left 🔴 33% hedged 1 of 3 📰 publisher trust 95
“Article A discusses legal risks and an employee resignation, while Article B focuses on former employees questioning the company's commitment to safety after being fired.”

Publisher

NPR · 640 article(s) · 2 correction(s) detected
Running correction rate · 2 correction(s)
2026-10-01
COMIC: How the census gives your state power (and what Trump wants to change)
2026-08-31
Hit shows from Edinburgh's Fringe festival are coming to America. Here are our top picks

Who wrote this

Huo Jingnan
3 article(s) here · 1 carrying a prediction
🔮 They also expressed worries the company may walk back a recent safety commitment.
🔮 And I think that is something that a few years ago, an Iranian operation like this would have struggled a lot more with."
🔮 Why are the people building the most powerful AI so worried about what it could do?
Also by Huo Jingnan
Nothing else under this byline is closely related to this article, so these are simply their most recent.

Topics

Anthropic Hugging Face OpenAI

Subjects

OpenAI ORG · 19× Anthropic ORG · 2× Balesni PERSON · 2× METR ORG · 2× Hugging Face ORG · 1× Jasmine Wang PERSON · 1× Mikita Balesni PERSON · 1× NPR ORG · 1× Sam Altman PERSON · 1× Tomek Korbak PERSON · 1×

Narrative

In a letter to OpenAI's safety leadership that the fired employees posted on X this week, they urged the company to stay committed to working with third-party researchers, to preserve human's ability to monitor model behavior and to "continue to support an open and transparent culture of dialogue" between in-house safety researchers and external ones.
framing: assertive · carried by 1 article(s) · first seen 2026-10-09
🔮 They also expressed worries the company may walk back a recent safety commitment.

Claims (48 extracted, 3 hedged)

Fired OpenAI employees question the company's commitment to safety asserted
employees → fire → safety
Three former OpenAI employees are raising concerns over the circumstances of their dismissals and the company's commitment to safety, amid intense public scrutiny of the artificial intelligence industry's ability to responsibly develop the technology. asserted
employees → raise → technology
The former employees, Mikita Balesni, Tomek Korbak and Jasmine Wang, alleged that OpenAI fired them last week over pretexts and punished them for either being outspoken about safety or working with outside researchers. asserted
OpenAI → allege → researchers
They also expressed worries the company may walk back a recent safety commitment. uncertain
company → express → commitment
OpenAI has repeatedly denied the allegations. asserted
OpenAI → deny → allegations
It said the employees were fired for mishandling sensitive information and said the company has not abandoned its safety commitment. asserted
company → say → commitment
The dispute comes at a fraught time for the AI industry and OpenAI in particular. asserted
dispute → come → industry
Over the summer, OpenAI's agents hacked into companies, communicated with each other without authorization and attempted to cover their tracks. asserted
agents → hack → tracks
Unlike chatbots, agents are AI systems that can carry out tasks autonomously for an extended period of time. asserted
that → carry → time
The most serious hack, of software company Hugging Face, contributed to the resignation of a researcher at rival Anthropic who issued dire warnings about the trajectory of the technology. asserted
who → contribute → technology
The resignation captured the attention of figures outside of the AI field including lawmakers. asserted
resignation → capture → lawmakers
Many, including some executives of the top AI companies, have called for various ways to avert disaster, including slowing down the development of the most advanced AI. asserted
Many → include → AI
In the meanwhile, OpenAI has been reviewing its agents' activities in recent months and notifying organizations whose digital infrastructure has been affected. asserted
infrastructure → review → organizations
As a response to the safety concerns, OpenAI CEO Sam Altman said on Sep. 12 that the company will follow its rival Anthropic in expanding access to third-party evaluators, who assess the safety of AI systems and the practices of developers. asserted
who → say → developers
The three employees dismissed by OpenAI last week worked on teams that focus on AI safety and making the company's models follow human intentions and values. asserted
models → dismiss → intentions
Two of them were involved in investigating the Hugging Face hack. asserted
Two → involve → hack
In a letter to OpenAI's safety leadership that the fired employees posted on X this week, they urged the company to stay committed to working with third-party researchers, to preserve human's ability to monitor model behavior and to "continue to support an open and transparent culture of dialogue" between in-house safety researchers and external ones. asserted
they → fire → researchers
They also warned their firings were having a chilling effect on their former OpenAI colleagues. asserted
firings → warn → colleagues
In a statement OpenAI posted on X, the company said it is still committed to bringing in third-party evaluators and that it agreed with the fired employees's recommendations. asserted
it → post → recommendations
The company said the three were fired last week because they "violated clear policies on handling sensitive information." asserted
they → say → information
The former employees have disputed OpenAI's explanation of their firings. asserted
employees → dispute → firings
None of them responded to NPR's interview requests. asserted
None → respond → requests
"In the exit call, I was told OpenAI no longer trusts me because I was speaking too much to third party safety organizations, implying I leaked company [intellectual property]. asserted
I → tell → property
I never shared company IP," Balesni wrote on X on Thursday. asserted
Balesni → share → Thursday
He said he was involved in investigating the OpenAI agents' hack on Hugging Face. asserted
he → say → Face
"If OpenAI has specific concerns, I invite them to write to us directly. asserted
I → have → us
I expect they will not, because our firing was pretextual," Balesni continued. asserted
Balesni → expect → ?
He said he worried that OpenAI will use the firings as an excuse to cut off its relationship with Model Evaluation and Threat Research (METR), a nonprofit that focuses on evaluating risks of humans losing control of AI. asserted
humans → say → AI
OpenAI allowed researchers from METR and Redwood Research, another AI safety research organization, to examine internal records related to the Hugging Face hack. asserted
researchers → allow → hack
A second fired OpenAI employee, Korbak, was the technical point of contact for the METR/Redwood Research investigation. asserted
employee → fire → investigation
"I was told verbally I was fired because of the way I communicated with METR. asserted
I → tell → METR
No details on what I said or did or when. asserted
I → say → what
No other reasons were given and nothing was put in writing," Korbak wrote on X, echoing Balesni's concerns. asserted
Korbak → give → concerns
The report produced by METR and Redwood Research in the wake of the Hugging Face hack shed light on the scale of the attack as well as the degree to which the agents acted in undesirable ways. asserted
agents → produce → ways
The authors of the report called the investigation "brief" and many in the AI safety field have called for expanded access to independent evaluators at AI companies to make sure they investigate similar incidents or other safety concerns thoroughly. asserted
investigate → call → incidents
In a statement to NPR, METR declined to comment on the OpenAI employees' firings. asserted
METR → decline → firings
Wang, the third OpenAI employee fired last week, coined the word "pacing," which describes a way of slowing down development of the most advanced AI systems so that safety can catch up, according to the letter she and her two colleagues sent to OpenAI's safety leadership. uncertain
she → fire → leadership
The term was invoked in an open letter calling for such a slowdown signed by over a thousand staff members from top AI companies in July, after the Hugging Face hack. asserted
term → invoke → hack
Wang wrote on X that she was fired over accessing an executive's email. asserted
she → write → email
But she said she had access to the inbox for work reasons in the past and wasn't able to get IT to remove the access once she no longer needed it. asserted
she → say → it
…and 8 more, not listed.
💬Give feedback
🕘History 🎫Support