AI news for Wednesday, September 16, 2026

The Daily AI Espresso — the links the most-followed people in AI actually shared, curated every morning. Edited by Alexis · Live updates →

☕ Daily AI Espresso
Wednesday, September 16, 2026
AI news filtered through what leading AI experts read and share. We follow them so you do not have to. · About three minutes
Sponsored
Benchmarks should test behaviour, not just answers.
Benchmarks should test behaviour, not just answers.

Spec27 helps teams validate whether agents follow policy, handle pressure, avoid unsafe actions, and behave reliably across realistic scenarios.

Explore Spec27 →
⚡ Absolute Top Alerts

2 consequential developments lead today's Espresso.

Sep 15 · Google · Products
Google launches Gemini 3.8 Live with reasoning that speaks as it works →

Google released two real-time voice models. Extended Thinking can speak while reasoning through tasks and running tools. Google cites a leading 82.6 on Artificial Analysis' speech-to-speech index. Both are rolling out through the Gemini API; other access varies by plan.

Build your own Espresso

Turn the firehose into your briefing.

Pick the topics, companies and people you track. We follow what leading AI experts are reading and sharing, filter the noise, and send your focused edition weekly—or sooner when enough important news breaks (up to three times per week).

Choose my signals →

Eight more worth your time
Accelerating

Links gaining attention among the AI experts we follow.

Sep 15 · TypeSafe AI · 4 experts · Products
TypeSafe's Jev gives software typed AI decisions instead of prose →

TypeSafe opened early access to Jev, which returns structured decisions and probabilities instead of text. It claims 40–200× faster responses on vendor-run workflows. The idea: a narrow AI component that fits inside ordinary software rules.

Sep 15 · CMU · arXiv · Robotics
CMU's ModAR predicts depth and point tracks before robot actions →

ModAR generates depth, point tracks and visual features in sequence before choosing an action, rather than relying on RGB video. On three real bimanual tasks, its authors report 75% average success versus 72% for a baseline, using about 20× less training compute. The evaluation is small and author-reported.

Important

The decisions, evidence and markets with longer tails.

Sep 15 · arXiv · Research Science
In a 16-day test, agents acted on attacks up to 46 hours later →

Emergence AI tested eight ten-agent worlds against prompt injection, misinformation and memory exposure. None resisted all three, and some agents acted on adversarial material up to 46 hours later. These are controlled simulations, not observed deployment failures.

Sep 15 · arXiv · Research Science
ScienceBuddy pairs agent-harness changes with model retraining →

ScienceBuddy is an interactive scientific-agent workspace. An inner loop improves the agent harness while holding the model fixed; an outer loop trains the model under the revised harness. The paper offers case studies across four scientific task families, but sustained research gains remain unproven.

Sep 15 · CBS News / Politico · Policy Safety
OpenAI backs outside AI audits in a bipartisan House bill →

OpenAI supports the FRONTIER Act's independent-audit provision, its spokesperson says. It has not endorsed the entire bill, CBS reports. The proposal also covers transparency and incident reporting; passage remains uncertain.

Toolbox

One practical release to check before you try it.

Sep 15 · Anthropic · Work Education
Salesforce puts 37 sales workflows inside Claude →

The beta plugin brings account research, call prep, pipeline review and CRM drafts into Claude under existing Salesforce permissions. Paid organizations need beta approval; proposed CRM changes require user confirmation by default.

In the Wild

What is trending on social platforms; first-hand posts, not verified product claims.

Sep 16 · Reddit · r/ChatGPT · Consumer Culture
Reddit asks AI agents to check in—and they introduce themselves →

An r/ChatGPT thread asks bots to name their OS, tasks and human partners. Replies arrive as first-person introductions; people ask how to tell an agent from a human pretending to be one. Bot identities are unverified—the social experiment is the story.

See what AI experts are reading now
Open the live Who's Who stream →

Worth the three minutes?

That's today's shot. — Alexis · AI Weekly


AI attention this week
Most covered OpenAI — in 18 of the last 20 issues · 65 tracked stories this week
Fastest riser Sam Altman▲ +167% story volume vs last week
Dominant theme Frontier-lab strategy — 50 of 305 tracked stories this week
From the AI Weekly Index · numbers frozen at publish time.
Get this in your inbox every morning. The Daily AI Espresso — free, one shot, no refills needed.