Anthropic finds frontier models can automate targeting and weapons software—and documents real misuse
Anthropic released evaluations showing frontier models performing parts of tactical intelligence and conventional-weapons engineering that historically required scarce expertise. Mythos-class models beat an elite-human proxy on outdoor-photo geolocation, while Opus 5 independently wrote and iterated simulated drone guidance that struck moving vehicles in 47% of easier trials and hit targets in 20% of 540 launches across all nine settings; the hardest camouflage, decoy, and GPS-spoofing conditions largely remained unsolved.
Why it made the cut: This is the first substantial evaluation suite connecting model scaling to intelligence targeting and conventional-weapons development, and it is paired with evidence of real use rather than simulations alone. Anthropic says it disrupted six weapons-related operations—including a Yemen-based group that used Claude for guided-rocket software and conducted a failed field test—and consequently deployed new weapons-development classifiers; the findings are self-reported and do not establish how much AI improved the actors’ outcomes.
Technical evaluation and official announcement · Threat-intelligence report (PDF) · Detailed threat report · Independent coverage (AP)
Link to this post