Anthropic blocks bioweapon research attempts in third AI misuse report
Anthropic has released its third report on AI misuse, detailing how its systems blocked malicious actors from using Claude for cyberattacks, surveillance, and biological weapons research. The report highlights a specific instance where the AI refused a request to draft a grant proposal for enhancing the chikungunya virus, a dual-use query that could lead to more dangerous pathogens. While older…
Key points
- Anthropic blocked requests to enhance chikungunya virus mutations, citing dual-use biological risks.
- Researcher Jacob Coxon resigned over concerns regarding the race to self-improving superintelligence.
- UK politicians urged a ban on ASI, but the government rejected it in favor of safety regulations.
The disclosure coincides with internal tensions at the company, as researcher Jacob Coxon resigned, citing concerns that Anthropic and OpenAI are racing toward self-improving superintelligence without adequate safety measures. This resignation aligns with broader political pressure in the UK, where over 70 MPs and peers urged Prime Minister Andy Burnham to ban the creation of artificial superintelligence. The UK government responded by rejecting a ban, emphasizing a science-led approach and the need for robust cyber defenses rather than restricting model access.
Anthropic stated that the reported cases represent novel and notable threats rather than typical misuse, including industrial-scale model distillation and state-sponsored propaganda operations. The company urged governments and competitors to adopt similar defensive measures, arguing that as AI capabilities grow, the responsibility to prevent misuse must be shared across the industry and public sector.
The story so far
4 episodes →- Anthropic blocks bioweapon research attempts in third AI misuse reportthis story
Claude owner blocked AI misuse that could have supported biological weapons
yahoo.com · 11 September 2026
Claude owner Anthropic says it has blocked efforts by bad actors to use its artificial intelligence models for malicious activity such as cyberattacks, surveillance and research that could have led to biological weapons.
As AI models grow more powerful, elaborate cyberattacks no longer require sophisticated skills, and even lone individuals can create threats that would not have been possible even a year ago, Anthropic said.
The company said it has added stronger safeguards in its latest models to restrict biological research that could also be used to make weapons.
"The cases we share here aren't typical misuse, but rather examples of the most notable and novel threat activity we've identified to date," Anthropic said in its third report since March 2025 describing AI misuse.
The report includes snippets of the malicious code and AI prompts Anthropic said it found, and it urged governments and AI competitors to identify and prevent similar abuse.
"We're publishing this work because we believe we have a responsibility to disclose malicious misuse of our services. As models become increasingly capable, their risks will increase, unless AI developers and society's defenders act to make them safer," the company said.
The lengthy report by the AI startup, which is planning an initial public offering this autumn, was published two days after one of its researchers announced he is resigning over concerns that Anthropic and its competitors are not acting responsibly in AI development.
He echoed concerns raised inside and outside of the industry about the technology's potential to elude human control.
In the UK, as a result of those warnings, more than 70 MPs and peers wrote a letter urging Prime Minister Andy Burnham to back a ban on the creation of artificial superintelligence (ASI).
The politicians, several of them former government ministers, urged Burnham to back a bill put forward in the Commons on Tuesday that would prohibit ASI – an as-yet-nonexistent AI that would outcompete human capabilities in most domains.
A government spokesperson said: "Britain cannot simply turn AI off, and blocking access to models in the UK would not prevent them being developed or misused elsewhere. Companies have a clear responsibility to develop their products safely and to invest in the security infrastructure this technology requires."We will continue to take a long-term, science-led approach to understand and prepare for emerging risks from AI – including through the work of our world-leading AI Security Institute. Our Cyber Security and Resilience Bill will boost UK cyber defences and improve the cyber security of our essential public and digital services that we all rely on, including data centres."
Between December 2025 and August 2026, researchers at Anthropic found misuse by actors ranging from spyware vendors and "politically motivated individuals" to state-sponsored groups spreading propaganda.
Among the findings in the company's report are unnamed actors attempting to use its models for research that could have led to biological weapons.
In one instance, Anthropic said its systems blocked a request for Claude's assistance in authoring a grant application for scientific funding on the chikungunya virus.
Chikungunya is a mosquito-borne virus that causes debilitating symptoms such as severe pain and fever. The request involved a grant proposal for research seeking to enhance mutations to make the virus progressively more harmful. While such research could "certainly" be used to develop better vaccines and treatments, Anthropic said, "it could also be used to make the pathogen more dangerous".
None of the cases Anthropic included in its report was found to be using its newer, more powerful Claude Fable or Mythos-class models, with the exception of one illicit distillation case that Anthropic described as "an industrial-scale, covert campaign to extract a model's capabilities and replicate them in another model without authorisation."
Anthropic said its older models, such as Claude Opus 4 and Claude Sonnet 4.5, from 2025, "were well below the threshold where they could meaningfully assist a sophisticated user in carrying out dangerous biological research."
"As a result, safeguards on these models were less stringent, directed mostly at preventing access to content that might uplift novices in recreating known bioweapons," the report said. "But for today's models — which are capable of assisting in a range of complex scientific research tasks — the evidence is no longer certain, and we cannot make that same assurance."
Because of this, Anthropic has applied "stronger safeguards that restrict access to a wide range of dual-use biological research queries" in its more recent models, such as Claude Fable 5, the report said.
As companies introduce increasingly powerful AI models, experts have called on governments to regulate the technology, rather than relying on the industry to police itself.
Anthropic also found groups that created hundreds of social media accounts that look like they belong to ordinary people and then posted material amplifying the same political view over the course of a week. The company outlined nine such cases it found, originating in Russia, Iran, Turkey and across the Persian Gulf, South Asia, Africa and Europe.
While social media companies can detect influence operations on their platforms once posts are circulating, "we may see it on Claude while the operation is still being built".
Anthropic released this report after one of its researchers, Jacob Coxon, announced he's resigning amid fears the company and its chief rival OpenAI "are racing straight to self-improving superintelligence and gambling with our lives." Coxon's post warned that some of his colleagues now believe AI could threaten human life by the end of the decade.
But Anthropic said it has blocked each of the malicious activities it identified, used the experience to strengthen safeguards and shared information with government authorities and industry partners.
"We hope that the findings in this report will help other developers recognise similar patterns on their own platforms, give governments and civil society a clearer view of how emerging threats take shape, and strengthen collective defences," Anthropic said.
Reporting History sees journalists join News At Ten anchor Tom Bradby to revisit their remarkable on-the-day reports of the defining events of the modern age. Listen to the episodes below...
This text was published by yahoo.com and written by ITV News. It is reproduced here with attribution so you can read it in full; the rights remain with the publisher. Read it at the source ↗
Coverage and discussion
3sources- Anthropic Has Resumed External Cybersecurity Testing. Is That Enough?Press · nationalinterest.org ·
- Anthropic Says It Blocked Claude Users From Researching Biological Weapons. They Sought To Evade Restrictions.Press · ibtimes.com ·
The headline, key points and digest above were generated by Digest AI's editorial model from the linked sources. Automated summaries can contain errors: the sources are the record. Spotted a mistake? Tell us.
More in Policy & Regulation
All →- FAA may launch $875 million AI program Smart to aid air traffic control, WSJ reports · 1 src
- King Charles urges AI leaders to establish controls before it’s too late · 4 src
- OpenAI reveals six new AI misalignment incidents and launches reporting framework · 31 src
- Executives debate whether AI safety concerns mask a push for control · 2 src
- Antitrust experts warn AI slowdown proposals could trigger investigations · 1 src
Comments
via GitHub Discussions