Research worth keeping up with

AI Research Newsletter.

Consequential AI research and releases, with the context that makes them matter.

A curated digest of work from across the field. Prepared with AI assistance and primary-source links; these are summaries of others’ research.

Daily · Up to three updatesSignal over volume

Misuse

← All updates · 1 update

Anthropic finds frontier models can automate targeting and weapons software—and documents real misuse

Anthropic released evaluations showing frontier models performing parts of tactical intelligence and conventional-weapons engineering that historically required scarce expertise. Mythos-class models beat an elite-human proxy on outdoor-photo geolocation, while Opus 5 independently wrote and iterated simulated drone guidance that struck moving vehicles in 47% of easier trials and hit targets in 20% of 540 launches across all nine settings; the hardest camouflage, decoy, and GPS-spoofing conditions largely remained unsolved.

Why it made the cut: This is the first substantial evaluation suite connecting model scaling to intelligence targeting and conventional-weapons development, and it is paired with evidence of real use rather than simulations alone. Anthropic says it disrupted six weapons-related operations—including a Yemen-based group that used Claude for guided-rocket software and conducted a failed field test—and consequently deployed new weapons-development classifiers; the findings are self-reported and do not establish how much AI improved the actors’ outcomes.

Technical evaluation and official announcement · Threat-intelligence report (PDF) · Detailed threat report · Independent coverage (AP)