Research worth keeping up with

AI Research Newsletter.

Consequential AI research and releases, with the context that makes them matter.

A curated digest of work from across the field. Prepared with AI assistance and primary-source links; these are summaries of others’ research.

Daily · Up to three updatesSignal over volume

September 28, 2026

← All updates · 2 updates

NVIDIA releases OpenShell 0.1.0 with formal permission checks and a separate hardware watchdog design

NVIDIA launched an agent-safety platform combining its open-source OpenShell runtime with Sentry, a reference design for monitoring and enforcement on separate BlueField-4 hardware. OpenShell 0.1.0 adds formal policy analysis, protected credentials, and controls over individual API operations; NVIDIA reports that combined review and runtime controls prevented protected-repository writes in adversarial tests lasting up to two hours.

Why it made the cut: Released code and a documented enforcement architecture give builders concrete tools for containing agents outside their own reasoning and tool harnesses. The policy proofs cover modeled permissions, not every implementation flaw or harmful action; the tests are vendor-reported, and Sentry's millisecond-quarantine claim is not an independently validated containment guarantee.

Technical walkthrough and experiment summary · Official platform announcement and architecture · OpenShell code

Sonnet 5.5 reports a large terminal-task gain and brings stronger cyber safeguards to the Sonnet tier

Anthropic released Sonnet 5.5, reporting 70.6% on Terminal-Bench 4.0 versus Sonnet 5's 10.3%, alongside lower token use and faster generation. Its cybersecurity capability is now comparable to Opus 5's, prompting the first Sonnet launch with advanced cyber safeguards and fallback to an older model for higher-risk requests.

Why it made the cut: The size of the reported agentic gain and the changed safety requirements make this more consequential than a routine speed update. These are Anthropic's reported results, with performance dependent on effort settings and evaluation setup; the company says Sonnet 5.5 does not advance its overall capability frontier and that Opus 5.5 remains stronger on complex, open-ended work. The linked system card could not be retrieved for this briefing, so detailed safety conclusions remain unverified here.

Official release, evaluation table, and safeguards