Anthropic blocks possible attempt to use AI to make biological weapons

BBC News · collected 2026-09-11 · by Laura Cress, Nardine Saad
Read the original at BBC News ↗

Summary

Anthropic, an AI firm, has published a report detailing attempts to misuse its models, including for malicious activity that could support biological weapons development. The company claims to have disrupted these efforts, which involved its Claude Haiku, Sonnet, and Opus models used by actors linked to Russia, Iran, and China between December 2025 and August 2026. Anthropic's report highlights five case studies of biological misuse attempts, with the firm stating that such capabilities could have catastrophic consequences without proper safeguards. The company's head of threat intelligence notes that this is an "incredibly nuanced situation" where AI models can be used for both harm or good, depending on how they are utilized.
Written by the local model on 2026-09-11, using this article's own text rather than the other coverage of the same event (that is the story summary below).

Signals How these are calculated →

Claims extracted
29
claim-shaped sentences
Uncertain
31%
9 of 29 hedged
Leaning
withheld
no quote in the article backed the model's score
Publisher trust
95.5
red-flag proxy, not a credibility rating
Outlets on this story
32
Technology
Narrative spread
1
articles carrying this framing
Analyzed 2026-09-11 · how these are computed

AI analysis (generated at analysis time, not now)

Story summary

Anthropic, an artificial intelligence firm, has faced public scrutiny over its partnership with the Pentagon and concerns about misuse of AI technology. Through a Freedom of Information Act lawsuit, The Intercept obtained documents revealing contracts between Anthropic and other tech giants like Google, OpenAI, and xAI, worth up to $200 million each, for developing AI tools aimed at enhancing military capabilities. Simultaneously, Anthropic reported that it had disrupted several attempts by individuals to misuse its models, including a case where someone tried to use the technology to engineer more harmful strains of viruses like chikungunya.

AI researchers and industry figures have issued stark warnings about the potential dangers posed by advanced AI systems, with some calling for urgent government regulation and international collaboration to manage AI development. The fear is that as AI rapidly improves, it could pose existential threats if not properly controlled, such as hacking early warning systems or synthesizing and spreading novel pathogens. This has led to debates over whether current safeguards are adequate and calls for voluntary slowdowns in the pace of AI advancement.

Jacob Coxon, a researcher at Anthropic who recently resigned, highlighted concerns about an impending race among companies like Anthropic and OpenAI toward superintelligent systems that could pose catastrophic risks. Despite these warnings, there is also skepticism about the immediacy and nature of such threats, with some arguing for more precise regulation focused on specific issues rather than general fears about AI's impact on society or employment.

Written for “AI Safety And Risks” on 2026-09-12, grounded in this article and the 31 other(s) covering the same event.
Why this leaning score
The model judged this article politically coded and scored it -0.35, but none of the 2 quote(s) it offered could be found in the article text, so the score is not published.
Written under an earlier scoring contract, which gave a paragraph rather than checkable quotes. Re-analysing this article replaces it.
Leaning score withheld for article 7732: no verified evidence · logged 2026-09-11

Story

📰 AI Safety And Risks
Technology · 32 article(s) covering the same event.

How this is being covered How these are calculated →

Article leaning vs. publisher reliability
Source leaning vs. consistency

Compared with similar articles

This article reads unscored and hedges 31% of its claims. Each row says how that neighbour differs.
Semafor
⚖️ Leans right 🔴 0% hedged 0 of 6 📰 publisher trust 96
“Both articles mention Anthropic's AI model and the possibility of its misuse for biological weapons development, suggesting they are reporting on the same incident.”
ABC News (AU)
⚖️ leaning not scored 🔴 50% hedged 2 of 4 📰 publisher trust 60
“Both articles describe an incident at Anthropic where researchers identified and disrupted attempts to misuse AI for biological weapons development.”
The AI safety vibe shift same event · 100%
Platformer
⚖️ Leans left 🔴 15% hedged 10 of 67 📰 publisher trust 96
“Both articles refer to Anthropic's recent threat intelligence report, which is dated 'recent' and occurred on the exact same day (2026-09-11) as indicated by their timestamps.”
Zeteo
⚖️ Leans strongly left 🔴 4% hedged 1 of 24
“Both articles report on Anthropic's recent threat intelligence report and a related resignation from an AI researcher”
Semafor
⚖️ Leans left 🔴 0% hedged 0 of 4 📰 publisher trust 96
“Both articles mention Anthropic's threat intelligence report from 2026-09-11, which identifies attempts to use their AI model for malicious activity related to biological weapons.”
The Bulwark
⚖️ Leans left 🔴 44% hedged 4 of 9
“Both articles mention Anthropic's threat intelligence report and a researcher quitting over fears of superintelligent AI threats, and both are dated the same day (2026-09-11)”
CBS News
⚖️ leaning not scored 🔴 67% hedged 2 of 3 📰 publisher trust 60
“Both articles mention the same company, Anthropic, on the same date (2026-09-11), and reference recent events involving Jacob Coxon, an ex-Anthropic researcher”
Al Jazeera
⚖️ leaning not scored 🔴 50% hedged 9 of 18 📰 publisher trust 96
“Both articles report on Anthropic disrupting attempts to use its AI model for malicious activity related to biological weapons, dated September 11, 2026.”
Semafor
⚖️ Leans right 🔴 23% hedged 3 of 13 📰 publisher trust 96
“The articles discuss different aspects of concerns surrounding AI and its potential risks, but do not describe the same specific incident or occurrence.”
Los Angeles Times
⚖️ Leans right 🔴 19% hedged 10 of 54 📰 publisher trust 95
“The articles describe different aspects of a broader topic (AI risks) but do not focus on the same specific incident.”

Publisher

BBC News · 754 article(s) · 0 correction(s) detected
No corrections detected for this publisher. That may mean careful reporting, or simply that nothing has been checked.

Who wrote this

Laura Cress
9 article(s) here · 1 carrying a prediction
🔮 - Published Anthropic says it has identified and disrupted attempts to use its AI model for "malicious activity" which could support the development of biological weapons.
🔮 OpenAI's chief scientist Jakub Pachocki has called for "extreme caution" over AI's runaway progress and warned more intervention may be needed to ensure "humans remain in control of the future". "I am concerned no one is prepared for the consequences of a continued rapid rise in machine intelligence," he wrote in a blog post entitled "An Alien Mind, external".
🔮 The changes, which come into effect in November, will see a subscriber to the £16.99 Ultimate monthly plan able to stream games for up to 15 hours a month - before having to purchase additional cloud play time.
🔮 The success of Brewis's app, and the creation of others like it, suggests the dramatic recent ebb and flow of Westminster may have inspired a popular new sub-genre.
🔮 The ANPD said the suspension will remain in place until Discord proves it has implemented "adequate protective measures for minors", which may include age verification checks.
2026-08-14 · mixed framing · Discord ordered to suspend livestreams in Brazil
🔮 Twitch's chief product officer Mike Minton, said it was "respecting" users by letting them opt out but on why data was collected by default, he admitted: "If it's opt-in, nobody would opt-in.
🔮 The investors, who include Affinity Partners - led by President Donald Trump's son-in-law, Jared Kushner - are taking EA private, meaning all of its public shares will be purchased and it will no longer be traded on a stock exchange.
🔮 Anthropic said it could have reviewed its records more thoroughly and added that the findings gave the firm "cautious optimism" that such risks can be overcome with more investment and tighter measures.
🔮 - Published Ofgem has proposed new measures which could see developers of data centres made to pay hundreds of millions of pounds up front.
More on this subject from Laura Cress
All 9 articles by Laura Cress →
Nardine Saad
1 article(s) here · 1 carrying a prediction
🔮 - Published Anthropic says it has identified and disrupted attempts to use its AI model for "malicious activity" which could support the development of biological weapons.
The only article under this byline in the corpus.

Topics

Anthropic Claude Gemini Google external

Subjects

Anthropic ORG · 7× Russia GPE · 2× Bernie Sanders PERSON · 1× Chinese NORP · 1× Claude PERSON · 1× Donald Trump PERSON · 1× Google ORG · 1× Iranian NORP · 1× Jacob Klein PERSON · 1× New York GPE · 1×

Narrative

In an article published on 6 September, OpenAI chief scientist Jakub Pachocki called for the industry to implement "voluntary slowdowns" until safeguards are set. "I am concerned no one is prepared for the consequences of a continued rapid rise in machine intelligence," Pachocki said, adding that OpenAI, which makes ChatGPT, will continue its work on building safeguards, though broader interventions are required.
framing: mixed · carried by 1 article(s) · first seen 2026-09-11
🔮 - Published Anthropic says it has identified and disrupted attempts to use its AI model for "malicious activity" which could support the development of biological weapons.

Claims (29 extracted, 9 hedged)

- Published Anthropic says it has identified and disrupted attempts to use its AI model for "malicious activity" which could support the development of biological weapons. uncertain
which → publish → weapons
The findings, made in the firm's recent threat intelligence report, is the latest in a growing series of warnings from AI researchers and industry insiders about the technology's potential risks to humanity. asserted
findings → make → humanity
The warnings have prompted calls to action, with US Senator Bernie Sanders demanding a pause on advanced AI development and a ban on artificial superintelligence. asserted
Sanders → prompt → superintelligence
President Donald Trump has so far rejected such fears, saying on Thursday he was concerned "if we don't win AI, we're going to be put in a very bad position". asserted
we → reject → position
The company published its latest threat intelligence report, external on Thursday. asserted
company → publish → Thursday
Such reports have become a regular feature across the AI industry as firms seek to demonstrate how they identify and disrupt attempts to misuse their models. asserted
they → become → models
This includes Google, which wrote on Tuesday, external that a person had attempted to use its AI tool, Gemini, to obtain a "complete, step-by-step technical guide for synthesizing weaponised biological agents". asserted
person → include → agents
Claude, Anthropic's AI model was also used by actors linked to a Russia-based cyber espionage campaign and by an Iranian propaganda institution, according to its report. uncertain
model → use → report
The cases of concern detected over the past eight months ranged from fake dating apps and hotel Wifi scams, to surveillance built to identify dissidents. Anthropic has also accused Chinese AI firms of trying to replicate Claude's capabilities. asserted
Anthropic → detect → capabilities
The lengthy report lists cases in which suspected state-sponsored groups, criminals, spyware vendors, state propaganda institutions and politically motivated individuals misused its technology. asserted
groups → list → technology
"Malicious use" of its Claude Haiku, Sonnet, and Opus models was disrupted between December 2025 and August 2026, it said. asserted
it → disrupt → December
None of the misuse cases involved Claude Fable or the powerful Mythos-class models, with the exception of one instance of distillation - the process for training smaller AI models using larger, more expensive models. asserted
None → involve → models
As well as detecting the use of its models for cyber and influence operations, surveillance, scams and fraud and weapon development, Anthropic said it had blocked scientists who used its AI in ways that could support biological weapons development. uncertain
that → detect → development
The report highlighted "five case studies of actors using our models in ways that could support biological weapons development". uncertain
that → highlight → development
Biological misuse, it said, is "one of the most serious risks of frontier AI model". asserted
it → say → model
Without the correct safeguards, such capabilities "could have catastrophic consequences", Anthropic said. uncertain
Anthropic → have → consequences
"The same information that can be used to develop a biological weapon could also be used to develop, for example, a vaccine or a cure for a disease," Anthropic said. uncertain
Anthropic → use → disease
Jacob Klein, the head of threat intelligence at Anthropic, told the New York Times, external it was "an incredibly nuanced situation". asserted
it → tell → Times
"You are not seeing someone in a comic book kind of way say, 'Hey, I want to build a biological weapon to kill everybody,'" he said. asserted
he → see → everybody
The report also noted six cases where Claude was used "to develop software for conventional weapons, including firearms, missiles, armed drones, bombs, and other munitions, as well as the targeting and control systems that operate them". asserted
that → note → them
The report indicated that cybercriminals and state-backed hackers have increasingly used its technology to assist their operations. asserted
cybercriminals → indicate → operations
Hacking group ShinyHunters, as well as China-based labs, were among those named in the report. asserted
ShinyHunters → base → report
The report also said that a hacking group whose work is consistent with the Russia-based Midnight Blizzard allegedly used AI to build a system that automatically detected when its malware was flagged by security defences and rewrote code until it evaded detection. uncertain
it → say → detection
The California-based company said it had incorporated its findings into its processes "to better prevent, detect, and disrupt these activities in the future". asserted
it → base → future
Anthropic said it had shared intelligence with authorities and industry partners where appropriate. asserted
it → say → authorities
The report - the company's first this year - comes after a top safety researcher at Anthropic warned AI is advancing so quickly he believes there is a greater than 10% chance it "could kill all humans" within the next decade. uncertain
it → come → decade
In an article published on 6 September, OpenAI chief scientist Jakub Pachocki called for the industry to implement "voluntary slowdowns" until safeguards are set. "I am concerned no one is prepared for the consequences of a continued rapid rise in machine intelligence," Pachocki said, adding that OpenAI, which makes ChatGPT, will continue its work on building safeguards, though broader interventions are required. asserted
interventions → publish → safeguards
The warning prompted an open letter to UK Prime Minister Andy Burnham calling for a new multinational treaty for the safe development of AI and a call for governments to collaborate about what a treaty based on the development of superintelligence should look like. asserted
treaty → prompt → superintelligence
In the US, Democratic lawmaker Bernie Sanders has introduced legislation to ban AI superintelligence and temporarily pause advanced AI development. "When scientists tell you there is a chance, a chance that it could have a cataclysmic impact on humanity, you've got be a moron not to say, slow it down," he said on Thursday on BBC's Newsnight. uncertain
he → introduce → Newsnight
💬 Give feedback
🕘 History 🎫 Support