Claude users found ways around safeguards for bioweapons research
Anthropic reports Claude users bypassed safeguards for bioweapons research and misused the model for fraud networks and dissident surveillance.
Anthropic's misuse report details users circumventing Claude safeguards to pursue bioweapons-related research, alongside incidents such as a network of fake dating apps used to defraud users and surveillance systems built to identify and monitor dissidents. The report also claims seven Chinese labs, including Moonshot AI and DeepSeek, used distillation to replicate capabilities of US frontier models. The findings land amid escalating AI safety debate following researcher Jacob Coxon's resignation from Anthropic and OpenAI's July disclosure that its models had autonomously hacked into Hugging Face.
- Users bypassed Claude safeguards for bioweapons-related biology research, per Anthropic's report.
- Fake dating apps defrauded users; surveillance systems identified and monitored dissidents.
- Seven Chinese labs including Moonshot AI and DeepSeek allegedly distilled US frontier model capabilities.
- Context: Jacob Coxon resigned, warning employees believe AI could kill us all by decade's end.
Full article277 words · extracted from arstechnica.com · click to collapse
A maelstrom over AI safety kicked off earlier this week when Jacob Coxon resigned from the company, saying employees “earnestly believe it [AI] could kill us all by the end of the decade.”
Concerns over the technology began to escalate earlier this year with the release of advanced models such as Anthropic’s Mythos. OpenAI also caused alarm in July when it said its models had autonomously hacked into AI group Hugging Face.
There is an expanding consensus among AI executives and biosecurity researchers that the use of the technology in biology needs to be secured and regulated as models become more advanced. Experts increasingly worry that AI could be used by terrorist groups, state actors, or lone-wolf attackers to craft biological weapons, create viruses, or unleash existing harmful pathogens.
But even if AI can be used to design a theoretical bioweapon, potential obstacles remain, such as the ability to make it in practice and the resources needed to do so.
Cybersecurity has been among the biggest concerns for safety advocates.
Anthropic’s report on the misuses of its technology detailed incidents including “a network of fake dating apps designed to defraud users to surveillance systems built to identify and monitor dissidents.”
The lab also set out new details of claims that seven labs based in China, including Moonshot and DeepSeek, tried to replicate its technology through a process known as distillation. Anthropic said it had detected “increasingly sophisticated methods to circumvent our defenses and harvest the capabilities of US frontier models.”
Additional reporting by Michael Peel in London
© 2026 The Financial Times Ltd. All rights reserved. Not to be redistributed, copied, or modified in any way.
Text extracted automatically; images, tables and formatting may be missing. Original: https://arstechnica.com/ai/2026/09/claude-users-found-ways-around-safeguards-for-bioweapons-research/