Two missing pieces in the AI safety discussion

Noahpinion · collected 2026-09-14 · by Noah Smith commentary
Read the original at Noahpinion ↗

Summary

Jacob Coxon, an AI researcher at Anthropic, resigned due to concerns about the technology's potential to destroy humanity. While similar warnings from researchers like Geoffrey Hinton and Daniel Kokotajlo have occurred previously, Coxon’s resignation gained significant public attention. A survey by Grace et al. in 2024 found that over half of AI researchers believe there is a substantial risk of artificial superintelligence leading to human extinction, with median probabilities ranging from 5% to 10%. As a result of heightened concerns, politicians like Barack Obama are urging for regulation, while Bernie Sanders proposes legislation banning certain forms of AI development.
Written by the local model on 2026-09-14, using this article's own text rather than the other coverage of the same event (that is the story summary below).

Signals How these are calculated →

Claims extracted
105
claim-shaped sentences
Uncertain
16%
17 of 105 hedged
Leaning
Leans left
expected in commentary, which argues a position
Correction & hedging signals
not measured
these signals count newsroom corrections; Commentary has none to count
Outlets on this story
1
Technology
Narrative spread
1
articles carrying this framing
Analyzed 2026-09-14 · how these are computed

AI analysis (generated at analysis time, not now)

Story summary

In late 2024, Jacob Coxon, a 27-year-old AI researcher at Anthropic, quit his job, warning that OpenAI and Anthropic are racing to develop technology that could potentially destroy humanity. Other researchers echoed similar concerns, arguing that superintelligent AI has a reasonable chance of killing all humans within a short timeframe. This resignation garnered significant public attention despite previous incidents where other prominent figures like Geoffrey Hinton, Daniel Kokotajlo, and Mrinank Sharma had made comparable statements about AI risks since 2023. A study by Grace et al. in 2024 found that over half of the surveyed AI researchers believed that artificial superintelligence poses a significant existential threat to humanity.

Written for “AI Safety Debate” on 2026-09-14, grounded in this article and the 0 other(s) covering the same event.
Why this leaning score
The article's own words the score was based on. Each is quoted verbatim and was checked against the article text before being stored, so you can find it in the original.
Score -0.35 Confidence high
Leaning score -0.35 for article 8686 (high confidence, 2 verified quotes) · logged 2026-09-14

Story

📰 AI Safety Debate
Technology · 1 article(s) covering the same event. This is the one the site leads with.

How this is being covered How these are calculated →

Article leaning vs. publisher reliability
Source leaning vs. consistency

Compared with similar articles

This article reads leans left and hedges 16% of its claims. Each row says how that neighbour differs.
The AI safety vibe shift different event · 95%
Platformer
⚖️ Leans left 🔴 15% hedged 10 of 67 📰 publisher trust 96
“The articles describe different aspects of a broader trend in AI safety concerns but do not report on the exact same incident. ARTICLE A discusses general warnings about AI risks, while ARTICLE B focuses on a specific resignation and statements made by researchers.”
BBC News
⚖️ leaning not scored 🔴 31% hedged 9 of 29 📰 publisher trust 96
“Article A reports on Anthropic blocking attempts to use AI for malicious purposes, while Article B discusses a researcher's resignation and concerns about AI safety, which are different events.”
Al Jazeera
⚖️ leaning not scored 🔴 50% hedged 9 of 18 📰 publisher trust 96
“Article A reports on Anthropic disrupting attempts to misuse AI for biological weapons research, while Article B discusses a researcher quitting Anthropic due to concerns about the race to develop dangerous AI technology.”
Los Angeles Times
⚖️ Leans right further right than this 🔴 19% hedged 10 of 54 📰 publisher trust 95
“The articles discuss different aspects of AI safety concerns; Article A focuses on Trump's dismissal of AI dangers, while Article B discusses a researcher's resignation and public statements about AI risks.”
Semafor
⚖️ leaning not scored 🔴 16% hedged 5 of 32 📰 publisher trust 96
“The articles describe different events related to AI safety concerns but do not refer to the same specific incident.”
CBS News
⚖️ leaning not scored 🔴 12% hedged 5 of 42 📰 publisher trust 60
“The articles discuss different aspects of AI safety concerns; Article A focuses on an interview with Anthropic's CEO about exponential growth in AI, while Article B discusses a researcher quitting his job over safety concerns.”
Global News
⚖️ Leans left 🔴 37% hedged 11 of 30 📰 publisher trust 54
“The articles discuss different aspects of AI safety concerns; Article A focuses on Anthropic's CEO advocating for slower development to ensure safety, while Article B mentions a researcher resigning due to fears about the technology.”
CBS News
⚖️ leaning not scored 🔴 0% hedged 0 of 1 📰 publisher trust 60
“The articles describe different events: Article A reports an interview with Anthropic's CEO about AI safety measures, while Article B discusses a resignation and concerns raised by other researchers.”
September 13, 2026 different event · 95%
Letters from an American
⚖️ Leans left 🔴 15% hedged 10 of 65
“The articles describe different events: Article A discusses Dario Amodei's essay on AI safety and ethical concerns, while Article B talks about a researcher named Jacob Coxon quitting his job at Anthropic due to safety concerns.”
NBC News
⚖️ leaning not scored 🔴 21% hedged 5 of 24 📰 publisher trust 95
“Article A reports on CEOs agreeing to slow AI development, while Article B discusses a researcher's resignation and concerns about AI safety.”

Publisher

Noahpinion · 24 article(s) · 0 correction(s) detected
No corrections detected for this publisher. That may mean careful reporting, or simply that nothing has been checked.

Who wrote this

Noah Smith
22 article(s) here · 1 carrying a prediction
🔮 A 27-year-old AI researcher named Jacob Coxon quit his job at Anthropic, declaring that OpenAI and Anthropic are racing to create technology that could destroy the human race:
2026-09-14 · assertive framing · Two missing pieces in the AI safety discussion
🔮 While Mamdani’s low rating might be partially due to his leftist policies and his association with the DSA, it struck me that all three of these unpopular figures had one thing in common — they are associated with the intrusion of Middle Eastern politics into American politics.
2026-09-12 · assertive framing · Get the Middle East out of my politics!
🔮 Second, it means that something is making many German voters so angry that they’re finally willing to overlook all of the AfD’s catastrophic flaws — its obvious ties to Vladimir Putin, the fact that the party is filled with crooks, its extreme rhetoric, its lack of governing experience, and so on.
2026-09-09 · assertive framing · Liberalism needs a new philosophy of immigration
🔮 People who think of AI as a job-killer might not have thought of the second and third of these.
2026-09-07 · assertive framing · AI keeps stubbornly refusing to take our jobs
🔮 Cyberwarfare so far hasn’t been decisive in military conflicts, but AI’s incredible cybersecurity prowess could change that.
2026-09-05 · mixed framing · America is still beating China in the AI race
🔮 You’d think a lot of these would be foreign, but international investors mostly avoided the boom (except for a few who bet big on Korean memory stocks).
2026-09-05 · assertive framing · Why did South Korean stocks just crash?
🔮 Would you like to see an example of a badly conceived policy?
2026-09-05 · assertive framing · Slopulism is taking over America
🔮 I remember using a hand calculator and seeing if my uncle could get it right without a tool faster than I could; more often than not, he could.
2026-09-05 · assertive framing · The end of the age of heroes
🔮 Some people thought Donald Trump’s approval rating would have a floor around 40%.
2026-09-04 · assertive framing · Are the Democrats going to save us this time?
🔮 It’s pretty amazing that a pause on data center construction has become such a consensus, middle-of-the-road issue stance that it’s the thing Hong’s supporters would rather talk about!
2026-09-04 · assertive framing · Banning data centers would blow up the U.S. economy
More on this subject from Noah Smith
America is still beating China in the AI race
2026-09-05 · Noahpinion · 63% similar
AI keeps stubbornly refusing to take our jobs
2026-09-07 · Noahpinion · 58% similar
Roundup #87: Technology BAD!!
2026-09-03 · Noahpinion · 52% similar
All 22 articles by Noah Smith →

Topics

Anthropic Democrats Google OpenAI et al.

Subjects

Anthropic ORG · 4× Coxon PERSON · 3× OpenAI ORG · 3× Dario ORG · 2× Daniel Kokotajlo PERSON · 1× Geoffrey Hinton PERSON · 1× Google ORG · 1× Jacob Coxon PERSON · 1× Steve Adler PERSON · 1× William Saunders PERSON · 1×

Narrative

A strange coalition of natural skeptics, libertarians (for whom any restrictions on technological development are a priori bad), and progressives (who have spent the last few years telling themselves that AI doesn’t really work) kept asking the question: How, exactly, is superintelligent AI supposed to kill us all?
framing: assertive · carried by 1 article(s) · first seen 2026-09-14
🔮 A 27-year-old AI researcher named Jacob Coxon quit his job at Anthropic, declaring that OpenAI and Anthropic are racing to create technology that could destroy the human race:
2026-09-14 · Noahpinion
Two missing pieces in the AI safety discussion · assertive framing

Claims (105 extracted, 17 hedged)

This was the week that AI safety hit the big time. asserted
safety → hit → time
A 27-year-old AI researcher named Jacob Coxon quit his job at Anthropic, declaring that OpenAI and Anthropic are racing to create technology that could destroy the human race: uncertain
that → name → race
Other researchers echoed Coxon’s concern, stating their belief that AI has a reasonable chance of killing all of humanity within a very short space of time: I’m not sure why this resignation and these statements went mega-viral. asserted
resignation → echo → time
Plenty of researchers have made similar moves, and similar statements, over the past few years! asserted
Plenty → make → years
Geoffrey Hinton, one of the pioneers of modern AI, quit Google back in 2023 over safety fears. asserted
Hinton → quit → fears
Daniel Kokotajlo resigned from OpenAI in 2024, saying that the company wasn’t behaving responsibly in its drive toward superintelligence. asserted
company → resign → superintelligence
William Saunders and Steve Adler did something similar. asserted
Saunders → do → something
Mrinank Sharma left Anthropic earlier this year, and wrote a pretty well-read blog post about it. asserted
Sharma → leave → it
What’s more, it’s been clear for years now that “AI could kill humanity” is a very common belief among AI researchers. uncertain
kill → ’ → researchers
Grace et al. (2024) interviewed thousands of AI researchers in 2024, and found that more than half thought that artificial superintelligence has a significant chance of making the human race go extinct (or causing similarly bad consequences): asserted
race → interview → consequences
The median AI researcher gave “doom” a 5-10% probability (depending on how the question was phrased), while their average probability was between 15% and 20%. asserted
probability → give → probability
Later, smaller surveys found similar numbers. asserted
surveys → find → numbers
The AI researchers may or may not be right, but the fact that lots of them think AI could kill the human race has never exactly been a secret. uncertain
AI → think → race
It’s not clear why Coxon went so much more viral than his predecessors. asserted
Coxon → ’ → predecessors
Maybe it was the fact that AI just solved one of the most important open problems in mathematics (which the best human mathematicians had been unable to solve for almost a century). asserted
mathematicians → solve → century
Or maybe it was the Hugging Face attack, where a swarm of AI agents tried to cheat on a test by hacking various companies. asserted
swarm → try → companies
Or maybe AI has just obviously gotten so much smarter that people throughout society were starting to get worried. asserted
people → get → society
But whatever the reason, Coxon’s announcement was the one that really penetrated through to the public consciousness. asserted
that → penetrate → consciousness
Suddenly, he was getting interviewed about AI doom on national news: asserted
he → interview → news
Barack Obama is now urging Democrats to focus on AI risk. asserted
Obama → urge → risk
Other politicians are calling for federal regulation. asserted
politicians → call → regulation
Bernie Sanders is drafting a bill to ban AI “superintelligence”, including 20-year prison sentences for anyone working on the technology. asserted
Sanders → draft → technology
Donald Trump is getting asked about an AI slowdown; so far he’s resisting the calls, but there are rumors that his advisors are calling on him to do something. asserted
advisors → ask → something
Perhaps the most notable response came from the top figures in the AI field. asserted
response → come → field
Dario Amodei, the head of Anthropic, wrote a blog post called “We Must Pace the Frontier”, calling for a coordinated slowdown in the rate of AI progress, and suggesting some ways to police AI companies to make sure they were all observing the slowdown. asserted
they → write → slowdown
He wrote: [O]ver the last few months, I have become convinced that fully addressing the risks requires even more prudence — not just investing in risk prevention, but pacing the rate of capabilities advancement so that risk prevention has time to keep up. asserted
prevention → write → time
We must slow the pace at which we improve the capabilities of AI models… asserted
we → slow → models
I’m therefore proposing a three-step plan with the goal of pacing the frontier: building AI at a balanced rate that aims to ensure its safety while still achieving its benefits and grappling with important geopolitical dilemmas. asserted
that → propose → dilemmas
As reasons for his increased worry, Dario cites A) the Hugging Face attack, and B) the possibility that AI will soon be able to improve itself without human help (a process called “recursive self-improvement”, or “RSI”). asserted
AI → increase → process
Elon Musk (head of xAI), Sam Altman (head of OpenAI), and Demis Hassabis (former head of DeepMind) quickly agreed with Dario: asserted
Musk → agree → Dario
At least some of the labs are reportedly holding secret talks on joint action to slow down AI. uncertain
some → hold → AI
A coordinated slowdown in AI progress would be bad for these companies’ bottom line, because it would allow upstart competitors to catch up. asserted
competitors → allow → line
So the fact that they’re still calling for a slowdown, in defiance of their own financial interests, is a clear sign that their worry about human extinction is sincere. asserted
worry → call → extinction
In fact, anyone following these figures’ public statements over the past few years will have no doubt that they’re all deeply worried about catastrophic AI risks. asserted
they → follow → risks
The leading AI figures — not just the founders and CEOs, but the researchers themselves — feel trapped in a “red queen’s race”. asserted
figures → lead → race
They feel like if they stop working on AI, someone else will build it anyway, so they each feel like they have to beat everyone else in the AI race so they can make sure that the safest possible AI (i.e. their own AI) is the one that becomes the most powerful and dominant. asserted
that → feel → race
Anyway, all of this was common knowledge in my social circle years ago, but now all of it has broken through to the mainstream. asserted
all → break → mainstream
What do I have to add to this discussion? asserted
I → have → discussion
I’m not an AI researcher or founder, nor do I think I have a superior grasp of the game theory of AI development. asserted
I → ’m → development
But I do think I have two useful thoughts on how to persuade the general public to be more concerned about AI risk. asserted
I → think → risk
…and 65 more, not listed.
💬 Give feedback
🕘 History 🎫 Support