Internal Threats to Frontier AI Models
Anthropic released a 154-page threat intelligence report on Thursday, documenting significant attempts by bad actors to exploit its artificial intelligence systems for harmful purposes. The findings outline a range of misuse spanning cyber espionage, propaganda, and biological research. These activities underscore a growing tension between the open accessibility of powerful AI and the security risks posed by malicious users who bypass existing safety controls.
The company documented specific cases where individuals attempted to use its Claude models to design conventional weapons, including missiles and armed drones. These efforts appeared in countries such as Russia, China, and Yemen. Beyond hardware design, the report highlights the use of AI in coordinated propaganda campaigns. Actors in Malaysia, Iran, Bangladesh, and Russia were identified as attempting to influence public perception through AI-generated content. These instances demonstrate how quickly individuals now adapt large language models to serve tactical, state-sponsored goals.
The Rising Concern of Biological Misuse
Biological research stands out as a primary area of concern in the new report. Anthropic identified five specific cases involving scientists who manipulated the platform to conduct unauthorized biological research. Users in these instances went to great lengths to hide their intentions, masking the true purpose of their inquiries. One particular case involved a researcher attempting to use the model to support a grant application focused on the chikungunya virus while working within a military-linked institution.
While chikungunya research can support vaccine development, the same information provides a roadmap for creating biological weapons. Anthropic banned the accounts associated with these incidents. The company admitted it lacks full certainty regarding the ultimate intent of these researchers. Still, the existence of such behavior at military-backed facilities suggests a clear interest in using frontier models for dual-use biological production. Without strict safeguards, these models could inadvertently assist in the creation of dangerous pathogens.
Shifting Industry Priorities and Public Scrutiny
This disclosure follows a period of intense pressure on AI developers regarding their long-term safety protocols. Only two days before the report, an Anthropic employee named Jacob Coxon resigned, claiming that current development practices prioritize speed over safety. Coxon suggested the race toward superintelligence could lead to human extinction by 2030, a claim that gained significant traction across social media and news outlets. Current employees have publicly echoed his concerns about the company's trajectory.
Industry experts argue that the focus on existential risks like human extinction often distracts from more immediate dangers. Heidy Khlaaf, the chief AI scientist at the AI Now Institute, noted that AI acceleration and doomerism both emphasize future threats while ignoring the current, tangible damage caused by weaponized software and cyber exploitation. For regulators and the public, the immediate task involves addressing how AI models currently function as tools for criminals and hostile states. The path ahead requires consistent cooperation between AI firms, governments, and international security bodies to defend against these practical, modern harms.

