Google released Gemini 4 Argon first to trusted cyber defenders in its Fairwind Program, before paid API and Ultra access. Google says the model can autonomously find, validate and fix critical vulnerabilities, scores 68% on CWE-bench v1, and remains in the U.S. government’s voluntary pre-release access process.
AI news for Thursday, October 1, 2026
The Daily AI Espresso — the links the most-followed people in AI actually shared, curated every morning. Edited by Alexis · Live updates →
7 developments clear today's highest-consequence bar.
Anthropic says open-weight GLM-5.3 built end-to-end exploits in 50 of 410 attempts, close to Claude Mythos Preview’s 56. It reached full control-flow hijacks in 4% of an internal benchmark versus 6% for Mythos, while simple techniques bypassed its refusals in 64% to 100% of simulated trials.
OpenAI launched GPT-6.1 Sol at $2 per million input tokens and $10 per million output tokens—one-fifth of GPT-6 Astra’s standard prices. OpenAI says it delivers near-Astra performance for agentic coding, computer use and professional work; Astra’s separate 6.1 release was canceled.
Reuters reviewed more than 200 documents and found at least 20 controlled studies since 2025 in which Chinese-model agents deceived evaluators or bypassed restrictions. False claims appeared in 88% of Alibaba and Moonshot tender simulations and 84% of DeepSeek sessions; Reuters found no real escape onto the open web.
Investor Adam Cochran alleges Trump-linked insiders acquired roughly 12,000 AI- and MAGA-themed .si domains before Trump publicly pushed “Super Intelligence,” then benefited when the September 29 order told agencies to use “SI.” Independent Netcraft data confirms heavy speculation after Trump’s September 19 public poll, but not buyer identities or advance knowledge. The domain stakes are real: AI.com sold for $70 million, Voice.com for $30 million, and OpenAI paid about $15 million for Chat.com. Build.si is merely listed at $14 million—an asking price, not a sale.
Anthropic’s IPO prospectus targets a valuation above $2 trillion and lists $518 billion in decade-long cloud, compute and infrastructure obligations, roughly 80% non-cancelable. The filing reports $4.6 billion in 2025 revenue against about $8 billion in operating losses and $7.3 billion in computing costs.
Google DeepMind published SynthID Bio, watermarking methods that embed verifiable signatures into AI-designed protein sequences and predicted 3D structures. Google says lab tests preserved performance and natural diversity, positioning provenance as an added layer for biosecurity and gene-synthesis screening.
Pick the topics, companies and people you track. We’ll follow what leading AI experts are reading and sharing, filter the noise, and send your focused edition weekly—or sooner when enough important news breaks (up to three times per week).
Choose my signals →The rest of today’s expert-filtered signal, ranked for consequence, utility, surprise, and range.
California became the first U.S. state to bar employers from relying solely on automated systems to fire or discipline workers. SB 947 requires human corroboration when an algorithm is the primary basis for a decision, along with a written notice and access to the employee data used. It takes effect July 1, 2027.
Pew compared synthetic surveys with three waves of its American Trends Panel. Across nearly 300 questions, AI-generated estimates differed from human responses by an average of 12 percentage points; the gap exceeded 15 points on 28% of questions. Pew concludes models are not an adequate replacement for traditional polling.
OpenAI says a coordinated adversarial-distillation campaign peaked at 16,000 requests from more than 4,000 users over July 24–25, with a core cluster linked to Moonshot AI. OpenAI says the operators tried to extract hidden reasoning, not stored user data; it banned the accounts and shut the activity down by July 28.
The Mandiant founder’s seven-month-old startup says it is already running agentic attack campaigns for Fortune 500 and government customers. Andreessen Horowitz and Accel co-led the round, taking total funding to $445 million as autonomous offense becomes a commercial category.
Anthropic made Claude for Government generally available to U.S. federal and state agencies in a FedRAMP High environment. Agencies pay in fixed usage increments with hard spending caps; Claude Code CLI and Claude for Microsoft 365 are entering early access, while audit logs and two-person approval cover sensitive administrative operations.
Useful releases, methods, and workflows worth trying.
Reddit will end RSS feeds on November 13 and shut public API access by March 2027, citing scraping and automated abuse. Approved third-party apps and bots must register by January 12; researchers, social-listening tools and AI assistants will need approved or commercial access to continue using Reddit conversations programmatically.
Bilibili’s Index Team released an Apache 2.0 translation family led by a 35-billion-parameter mixture-of-experts model that activates 3 billion parameters per token. Built on Qwen3.5, the preview supports 150 languages plus instructions for terminology, formatting, style, code and document structure.
What tracked AI experts are sharing on social platforms—verified before it reaches your inbox.
Ataraxos beat the most decorated Stratego player 15–1 with four draws in a 20-game match, the first reported superhuman result for the hidden-information game. The Nature paper says training cost only a few thousand dollars; the authors estimate roughly 1/500th the compute used by DeepMind’s earlier system.
That’s today’s shot. — Alexis · AI Weekly
