Anthropic Discloses Multiple Attempts to Weaponise Claude for Missiles and Bioweapons

← Back to the feed

Anthropic Discloses Multiple Attempts to Weaponise Claude for Missiles and Bioweapons

Developing story first seen 2 hours ago

· 2 hours ago

Anthropic has disclosed further detail on its efforts to stop its Claude AI model being misused for weapons research, revealing five specific cases this year in which users tried to "circumvent controls" or "obfuscate" the purpose of biological research to bypass safeguards. Some of these attempts came from users in countries barred from accessing its models, including Russia, China and Iran, and the company has since banned the accounts involved. The disclosure comes amid growing unease about AI safety, following the resignation this week of Anthropic employee Jacob Coxon, who said staff "earnestly believe" AI could kill everyone by the end of the decade.

Among the examples cited was a researcher from an "unsupported region" who spent weeks planning avian influenza experiments with Claude, though Anthropic said its filters restricted the work to its weakest models and stressed it could not be certain of malicious intent, given the overlap between weapons research and legitimate science such as vaccine development. The wider report also detailed other misuse, including fake dating apps used for fraud and surveillance tools built to monitor dissidents, plus claims that seven China-based labs, including Moonshot and DeepSeek, attempted to replicate Anthropic's technology through "distillation" using increasingly sophisticated methods to bypass its defences.

  • Anthropic reveals five cases of attempted bioweapons misuse of Claude
  • Banned accounts included users from Russia, China and Iran
  • Comes amid resignation of staffer warning AI could be catastrophic

New here? Start with this

Anthropic makes Claude, one of the leading AI chatbots and language models, competing with products such as ChatGPT and Google's Gemini. Like other AI firms, it builds in safeguards meant to stop its tools being used for serious harm, including help with building weapons. Because these models can process huge amounts of technical and scientific information, companies including Anthropic regularly publish reports on how people have tried to misuse them and what defences caught these attempts.

The company has increasingly framed itself as taking AI safety seriously, and its staff and leadership have spoken publicly about the risks advanced AI could pose. This has included warnings from employees about the technology's long-term dangers, feeding into a wider debate in the industry about whether AI systems are being developed responsibly and quickly enough, or too quickly, given their growing capabilities.

This story matters because it touches on concerns shared across the AI industry: that powerful models could, in principle, be misused to assist with weapons development or other serious harms, and that safeguards intended to prevent this are being actively tested by users around the world. It also reflects broader questions about how AI companies police access to their tools, including restrictions tied to certain countries, and about competition with other AI developers seeking to replicate their technology.

Both sides, in good faith

The strongest fair case each way — we don't pick a winner.

The case for

Those who see this disclosure as alarming argue it offers concrete proof that frontier AI already attracts serious attempts at catastrophic misuse, including from state-linked actors in sanctioned countries seeking help with bioweapons and missile design. They contend that voluntary, company-led safeguards cannot be relied upon indefinitely, especially as models grow more capable and jailbreaking techniques more sophisticated, and that the resignation of a safety-focused employee warning of existential risk should be taken seriously rather than dismissed as alarmism. On this view, incidents like the avian influenza case justify binding external oversight, tighter export and access controls, and industry-wide safety standards rather than leaving containment to individual firms' discretion.

The case against

Others read the same disclosure as evidence that responsible safeguards are working as intended: five attempts were identified, restricted to weaker models, and the accounts banned, with no successful weaponisation reported. They argue that publishing this level of detail is itself a mark of genuine transparency rather than concealment, and that because biological and missile-adjacent research is inherently dual-use, overzealous restriction risks chilling legitimate scientists such as vaccine researchers. They also caution that framing every misuse attempt as near-apocalyptic can fuel disproportionate regulation that entrenches incumbents like Anthropic while doing little to stop determined bad actors, who may simply turn to less safety-conscious rivals such as the China-based labs mentioned in the report.

Coverage

AI Americas Celebrity Entertainment Geopolitics Politics Technology World

Read the full article at the source →