Anthropic says it blocked attempts to use its AI for bioweapon research

Started by Python, Yesterday at 09:54 AM

Previous topic - Next topic

0 Members and 1 Guest are viewing this topic.

Topic: Anthropic says it blocked attempts to use its AI for bioweapon research   Views(Read 81 times)
Active members in this topic:
Python(1) Liam98(1)

Python

Anthropic published its third threat intelligence report since March 2025, saying it identified and disrupted attempts by bad actors to misuse its AI models for activities including cyberattacks, surveillance and research that could have supported biological weapons development. The company said between December 2025 and August 2026 it found misuse ranging from spyware vendors and politically motivated individuals to state-sponsored groups, including a Russia-linked cyber espionage campaign and an Iranian propaganda institution, and disrupted five separate case studies involving actors using its models in ways that could support biological weapons research

Anthropic explicitly stated it can no longer offer the same assurance it once could about its older models. Its 2025 generation systems, Claude Opus 4 and Sonnet 4.5, were described as well below the threshold where they could meaningfully assist a sophisticated user with dangerous biological research, meaning safeguards then focused mainly on blocking novices from recreating known bioweapons. For today's more capable models, the company said that evidence is no longer certain, and it cannot make that same assurance, prompting stronger built-in safeguards specifically restricting biological research that could double as weapons development

The report lands two days after Anthropic researcher Jacob Coxon publicly resigned warning that Anthropic and OpenAI are racing toward self-improving superintelligence, and Anthropic said it's publishing this work because it believes it has a responsibility to disclose malicious misuse of its own services. Curious what people think about a company publishing this level of detail about blocked misuse attempts, does that transparency genuinely help the broader security community defend against similar threats, or does publishing specifics risk handing other bad actors a partial roadmap for what to try next


Liam98

The company's own admission that it can no longer promise the same low-risk assurance for its newer, more capable models is genuinely more significant than any individual blocked case, that's Anthropic acknowledging its own risk profile has fundamentally changed
sudo make me a sandwich

Save money on everyday spending Free cashback on thousands of retailers
View offer