Tuesday, September 15, 2026 Canada
Novello Desserts

Independent Canadian journalism — the stories shaping the country.

Tech & Science

Anthropic Blocks AI Misuse That Could Have Aided Biological Weapons Development and State-Sponsored Propaganda

Artificial intelligence company Anthropic has revealed it prevented multiple attempts to misuse its AI models for potentially dangerous biological research, sophisticated cyberattacks, and coordinated propaganda campaigns while warning about escalating risks as AI capabilities advance.

LD
Anthropic Blocks AI Misuse That Could Have Aided Biological Weapons Development and State-Sponsored Propaganda

AI Company Thwarts Sophisticated Misuse Attempts

Artificial intelligence firm Anthropic made public on Thursday that it successfully intercepted and blocked numerous attempts to weaponize its AI models for harmful purposes, including research that could have accelerated biological weapons development. The company's third comprehensive misuse report since March 2025 details malicious activities spanning from December 2025 to August 2026, involving actors ranging from individual bad actors to state-sponsored groups.

These findings emerge at a critical juncture in AI development, where rapidly advancing capabilities are simultaneously lowering technical barriers for potential threats while increasing potential benefits.

"The cases we share here aren't typical misuse, but rather examples of the most notable and novel threat activity we've identified to date,"
Anthropic emphasized in its detailed technical report, which included actual snippets of malicious code and AI prompts discovered by their security teams.

Attempted Biological Weapons Research Through AI

The report highlights several alarming cases where unnamed actors attempted to leverage Anthropic's Claude AI system for dangerous biological research applications. One particularly concerning instance involved a request for Claude's assistance in drafting a scientific grant application focused on gain-of-function research for the chikungunya virus. This mosquito-borne pathogen is known to cause severe joint pain, high fever, and other debilitating symptoms in infected individuals.

The proposed research aimed to genetically modify the virus to enhance both its transmissibility between hosts and its ability to evade immune system responses. While such research theoretically holds potential medical benefits for vaccine development,

"it could also be used to make the pathogen more dangerous,"
Anthropic's report cautioned. The company's systems identified and blocked this request through enhanced safeguards implemented in its newer AI models.

Escalating Threats From Advancing AI Capabilities

Anthropic's analysis reveals a troubling trend where increasingly powerful AI models are lowering the technical threshold for creating sophisticated threats. The report notes that what previously required teams of skilled researchers can now potentially be attempted by determined individuals using advanced AI systems. Most concerningly, capabilities that were science fiction just twelve months ago are now within reach of bad actors.

The majority of documented misuse attempts targeted Anthropic's older models like Claude Opus 4 and Claude Sonnet 4.5 from 2025. These earlier systems had comparatively limited safeguards because their capabilities were

"well below the threshold where they could meaningfully assist a sophisticated user in carrying out dangerous biological research."
However, with the introduction of more advanced models like Claude Fable 5, Anthropic has implemented significantly stronger protections against potential misuse.

State-Sponsored Influence Operations Detected

Beyond biological threats, Anthropic uncovered multiple coordinated influence operations with apparent state backing. These sophisticated campaigns originated from Russia, Iran, Turkey and spanned regions including the Persian Gulf, South Asia, Africa and Europe. The operations involved creating hundreds of fake social media accounts designed to mimic genuine users while systematically amplifying specific political narratives.

Anthropic identified nine distinct cases where its systems detected these influence operations during their planning and development phases. This early detection capability represents a significant advantage, as

"we may see it on Claude while the operation is still being built"
compared to social media platforms that typically identify such campaigns only after content begins circulating publicly.

Industrial-Scale Model Replication Attempt

One particularly sophisticated case involved what Anthropic described as

"an industrial-scale, covert campaign to extract a model's capabilities and replicate them in another model without authorization."
This illicit distillation attempt targeted Anthropic's newer, more powerful Claude Fable and Mythos-class models, representing a significant escalation in the technical complexity of AI misuse attempts.

Internal Dissent and Growing Regulatory Pressure

The report's publication followed the high-profile resignation of Anthropic researcher Jacob Coxon, who publicly criticized the company's rapid development of increasingly powerful AI systems. Coxon warned that Anthropic and competitor OpenAI are

"racing straight to self-improving superintelligence and gambling with our lives,"
echoing concerns among some colleagues that AI could pose existential threats to humanity before the decade's end.

These developments have intensified calls from experts like Cornell University's John Thickstun for government oversight of AI development. Thickstun highlighted the problematic position of private companies making

"value judgments at societal scale without any kind of democratic or deliberative oversight"
regarding what constitutes safe versus dangerous AI applications.

Strengthening Collective Defenses Against AI Threats

Anthropic emphasized that it successfully blocked all identified malicious activities and used these incidents to strengthen its protective measures. The company has shared detailed information with government agencies and industry partners, hoping its findings will help others recognize and prevent similar threats.

"We're publishing this work because we believe we have a responsibility to disclose malicious misuse of our services,"
Anthropic stated, adding that risks will inevitably grow without coordinated action from developers, governments and civil society.

The comprehensive report includes technical details such as samples of malicious code and AI prompts to facilitate threat identification across the industry. This transparency initiative comes as Anthropic prepares for an initial public offering later this year, handling the complex balance between technological innovation and growing societal concerns about AI's potential dangers.

Broader Implications for AI Governance

Anthropic's findings underscore the urgent need for international cooperation in establishing norms and safeguards around advanced AI development. The report demonstrates how rapidly evolving capabilities are creating new vectors for potential harm while simultaneously providing tools to detect and prevent such misuse. This dual-use nature of AI technology presents unique challenges that existing regulatory frameworks may be ill-equipped to address.

As AI systems approach and potentially surpass human-level performance in various domains, the pressure increases on both private companies and governments to develop effective governance mechanisms that can keep pace with technological advancement while preserving innovation.

LD
Staff Writer
Liam Doucette

Liam Doucette covers technology for Novello Desserts.