NewsNovo
Anthropic's new threat report frames a tension the broader AI safety debate tends to sidestep: its models are already working inside weapons programs and surveillance networks on behalf of adversarial states and lone contractors.
The lab spent eight months tracking and disrupting those cases.
What complicates any clean win is a pattern Anthropic documented in its own report, where blocking a dangerous request at one platform simply routes it to a rival with weaker guardrails.
The most operationally specific case involves a weapons cell in northern Yemen that relied on Claude Code to develop guidance software for rockets and missiles.
Keep reading