Research worth keeping up with

AI Research Newsletter.

Consequential AI research and releases, with the context that makes them matter.

A curated digest of work from across the field. Prepared with AI assistance and primary-source links; these are summaries of others’ research.

Daily · Up to three updatesSignal over volume

Evaluation

← All updates · 1 update

Jev introduces fast, typed AI decisions, with early independent evidence for judging

TypeSafe AI released Jev on September 15, a specialized model that takes unstructured state and predefined questions and returns typed decisions with probabilities rather than generating prose. Vercel reported on September 18 that nearly 13% of its paid AI Gateway teams used Jev within its first 24 hours there; a September 22 independent preprint found Jev within three percentage points of its strongest LLM judge on preference and evidence-grounded factuality tasks at 0.36% of that judge's fee.

Why it made the cut: A model designed for low-latency classification, routing, and guardrail decisions could make a different class of AI-powered software practical, and both early platform use and an independent evaluation provide evidence beyond the launch claims. TypeSafe's larger speed and cost comparisons are self-run, Vercel's free introductory offer may have boosted early adoption, and the preprint reports larger gaps on tasks that require checking derivations or resisting elaborate wrong answers; none establishes general superiority or lasting production use.

Official announcement and technical discussion · Early adoption data (Vercel) · Independent evaluation (preprint)