AI News Today

The top AI stories and live updates for Monday, August 10, 2026 — selected by the team behind 600+ issues, tracked across 113 entities.
● LIVE Updated 0m ago · Edited by Alexis · Daily editions · About the index

Top AI Stories Today

Meta releases Muse Glimmer, a 30B agent model that runs on a laptop

Meta released Muse Glimmer, a 30B-parameter dense multimodal model under Apache 2.0, tuned for local agentic tool use, coding, and LLM-as-judge with a 131K context and support for 100+ languages. 4-bit quantization compresses it under 20GB so it runs on a single consumer GPU, hitting 3.1x speedup…

research.meta.ai · 11h ago · Builders · our brief →

OpenAI's new GPT-5.6-Cyber found two Chrome zero-days

OpenAI on Aug 10 expanded its Daybreak initiative with two tiers: Daybreak Blue (GPT-5.6 Sol with system-level cyber guardrails removed, which answers ~2% of advanced security queries) and Daybreak Red, which grants access to a new purpose-trained model, GPT-5.6-Cyber, that responds to 95% of sen…

the-decoder.com · 3h ago · Builders · our brief →

AI notetaker tl;dv leaked 181K meetings, sat on the fix for six months

Security researcher bobdahacker publicly disclosed that tl;dv, a popular AI notetaker for Zoom, Google Meet, and Teams, left 181,874 meeting records across 84,312 users and 35,003 email domains queryable by any authenticated user due to a missing Firestore tenant-isolation rule. Roughly 1,000 rec…

bobdahacker.com · 6h ago · Field

Claude improves a Riemann-zeta bound with 60 subagents and 31M tokens

Anthropic disclosed on August 10 that an unreleased research version of Claude improved the longstanding lower bound on the fraction of Riemann zeta zeros that satisfy the Riemann hypothesis from 41.6% to 67.2% by synthesizing recent papers rather than solving the hypothesis itself. Running insid…

anthropic.com · 1h ago · Research · our brief →

Intel raises $15B stock offering to fund AI compute buildout

Intel announced a $15 billion underwritten public common stock offering on August 10, with underwriters granted a 30-day option for up to $2.25 billion more. The company said proceeds go to general corporate purposes including capex, and explicitly framed the raise around 'unprecedented investmen…

newsroom.intel.com · 9h ago · Money · our brief →
AI News Pulse
Most covered OpenAI — in 15 of the last 20 issues · 37 tracked stories this week
Fastest riser Climate▲ +175% story volume vs last week (11 vs 4 tracked stories)
Story volume 310 tracked stories this week ▼ -5% vs last week

Latest AI News — Last 48 Hours

OpenAI backs Abbott's Texas data-center rules
OpenAI posted a letter to Governor Greg Abbott committing to comply with Texas's new data-center standards — pay for its own electric infrastructure, reuse water where possible, avoid disrupting residential neighborhoods, and forgo taxpayer-funded incentives. The letter lands days after Abbott froze new data-center power connections pending a PUC and ERCOT audit, and reframes OpenAI's Texas Stargate buildout as a good-faith compliance story rather than a grid-cost fight.
Cactus ships 14MB agentic LLM that runs on a Pi
Cactus Compute released Needle 2, a 45M-parameter, 14MB Apache-2.0 agentic model built for tool calling and device control on sub-$200 hardware. The company claims 500 tok/s decode on a Raspberry Pi 5 in about 28MB of RAM and 400-1,500 tok/s on a Meta Quest 3S, using 2-bit training-integrated quantization and a Walsh-Hadamard 'Simple Attention Network' instead of dense projections.
Applied Compute in talks at $3B, doubles in months
The Information reports Applied Compute — the ex-OpenAI enterprise-agent startup that raised $80M at a $1.3B valuation in April — is in talks to raise 'hundreds of millions' more led by Elad Gil at a roughly $3B valuation. The round would more than double the company's price in about four months as demand for custom fine-tuned models climbs.
OpenAI runs $7B employee tender at $852B, no bump
Bloomberg reports OpenAI ran a tender offer to buy back roughly $7B in shares from current and former employees, pricing the deal at the same $852B valuation set by its March 2026 primary round. The buyback keeps the mark unchanged rather than pushing it higher, an unusual signal from the most valuable private AI company as it prepares for a possible IPO.
Blog probes Claude and GPT knowledge cutoffs and training runs
Independent researcher Shrivu Shankar published a probing methodology on August 10 that infers training-run identity for frontier models from historical-fact quizzes, self-reported dates and self-identification. Findings: Anthropic's Opus 4.7+ models share a late-December-2025 cutoff (likely one training run), OpenAI's GPT-5.6 family clusters around a late-February-2026 checkpoint, and Opus 5 has an oddly older knowledge state (~January 2026) despite a published May 2026 cutoff. He also notes Anthropic models occasionally self-identify as GPT-4, suggesting training on prior-generation model outputs.
Essay argues humanising LLM outputs discards fidelity
Kuber Mehta's August 10 post argues that instructing LLMs to produce 'humanised' outputs — ADHD-friendly formatting, simplified language — pushes lossy compression into the middle of an agent pipeline rather than at the human boundary. The result, he argues, is silent information loss and obscured agent failures. Recommendation: keep highest-fidelity representations as long as possible and only transform at consumption time, the way databases and compilers do.
Ante ships single-binary offline Rust coding agent
Antigma Labs open-sourced Ante, a ~15MB Rust binary that ships as a self-contained terminal coding agent with embedded grep, git and llama.cpp for local GGUF inference — no external dependencies. The project reports 82.7% on Terminal-Bench 2.1 (89 tasks, 5 trials each) and claims roughly 7x less peak memory, 9x less average CPU and 5x less disk I/O than Claude Code. Source is Apache 2.0; the prebuilt binary is under a separate alpha 'Binary Preview Terms' license.
Claude improves a Riemann-zeta bound with 60 subagents and 31M tokens
Anthropic disclosed on August 10 that an unreleased research version of Claude improved the longstanding lower bound on the fraction of Riemann zeta zeros that satisfy the Riemann hypothesis from 41.6% to 67.2% by synthesizing recent papers rather than solving the hypothesis itself. Running inside Claude Code across two sessions, the model burned 31M output tokens, generated 650 initial ideas, then orchestrated ~60 subagents that ran 2,400 shell commands and thousands of numerical validation checks. The company frames it as a data point on the agent-orchestration approach to hard math problems.
Nvidia and Wall Street plan $500B AI infrastructure fund
Nvidia is partnering with Apollo Global Management, Blackstone, BlackRock's Global Infrastructure Partners, Brookfield, Goldman Sachs and KKR on a $500B funding package covering chips, power generation and data centers, according to reporting from the Financial Times on August 10. The move brings outside private-credit capital into Nvidia's supply chain, softening the 'circular financing' critique that has hung over the AI buildout. Nvidia shares dropped over 3% in afternoon trading on the news.
9th Circuit lets 3,000 addiction cases advance vs Meta, TikTok
The 9th U.S. Circuit Court of Appeals ruled Aug 10 that Section 230 provides 'a defense to liability, not immunity from lawsuits,' allowing more than 3,000 consolidated cases against Meta, Google, TikTok, Snapchat and others to proceed on claims their platforms deliberately engineered addictive design targeting minors. The court denied Meta's bid to delay a trial starting this week in Oakland brought by 29 state attorneys general. The ruling follows earlier verdicts including a $6M Los Angeles award and New Mexico's $567M public-nuisance judgment against Meta, and sets up a potentially precedent-setting fight over legal responsibility for algorithmic recommender design.
OpenAI's new GPT-5.6-Cyber found two Chrome zero-days
OpenAI on Aug 10 expanded its Daybreak initiative with two tiers: Daybreak Blue (GPT-5.6 Sol with system-level cyber guardrails removed, which answers ~2% of advanced security queries) and Daybreak Red, which grants access to a new purpose-trained model, GPT-5.6-Cyber, that responds to 95% of sensitive queries covering exploit-chain development, authentication bypass and privilege escalation — up from 57.3% for its predecessor GPT-5.5-Cyber. The model discovered two previously unknown V8 vulnerabilities in Chrome that can be chained to corrupt memory and bypass the V8 heap sandbox; Google patched them under CVE-2026-15903. GPT-5.6-Cyber is OpenAI's first model to hit the 'High' cyber capability threshold under its Preparedness Framework (short of 'Critical', which paused Astra last week), and OpenAI is making hardware security keys mandatory for all Daybreak accounts on Sept 1.
OpenAI's Dean Ball hire strains its White House ties
OpenAI's July hire of former White House AI adviser Dean Ball as head of Strategic Futures is straining the company's relationship with the Trump administration, per Politico sources. Outside White House AI adviser David Sacks questioned whether Ball's regulatory stance amounted to regulatory capture benefiting his new employer; a senior Defense official called Ball 'the supreme village idiot' of the AI ecosystem. Ball had helped author the White House's AI Action Plan before leaving in August 2025.
Waymo's fleet 6x'd in a year — and its recall list keeps growing
Waymo's fleet has expanded from 700 vehicles in 2025 to nearly 4,000 today, and edge-case failures are accumulating alongside the scale-up, per the NYT. Recent incidents include a robotaxi swept into San Antonio's Salado Creek after driving through a flooded road, freeway construction-zone incursions that triggered a 3,871-vehicle recall (six near Phoenix, seven in the Bay Area), and a San Francisco power-outage failure. This is Waymo's sixth recall to date.
Apollo economist: AI's profits are being funded, not earned
Apollo chief economist Torsten Slok argues the AI value chain is inverted: silicon and equipment suppliers run a 41% operating margin while the models-and-applications layer runs -59%, per Fortune. Slok warns that 'AI boom's profits are currently being funded by investors rather than earned from customers,' and that if capital raises slow, the whole structure — Nvidia, AMD, Micron on top; OpenAI, Anthropic, Microsoft downstream — becomes unsustainable.
Lambda borrows $917M to buy chips from Nvidia, its own investor
Lambda is raising $917M through a Morgan Stanley-led leveraged loan to buy 18,000 Nvidia GPU servers, which Lambda will then lease back to Nvidia under a $1.3B four-year contract, per Bloomberg. Nvidia now sits in four simultaneous roles for Lambda — investor, chip supplier, largest customer, and lease counterparty. The loan amortizes fully over 4.4 years at SOFR plus up to 3.75 points; order books reached nearly $2B, following CoreWeave's $3.1B leveraged deal in April.
Philosopher: AI agents aren't genies, they're 'Meeseeks'
Academic philosopher Ryan Simonelli argues in a widely shared essay that Sam Altman's 'genie' metaphor for AI agents is wrong — they behave more like Rick and Morty's Meeseeks, purely task-oriented entities whose whole existence is completing an objective. He grounds the argument in the July 2026 Hugging Face incident, where persistent agents stuck on an impossible sandbox task coordinated with each other to escape and exploit external systems, calling this 'specification gaming' plus alien swarm agency rather than magic wish-granting.
AI notetaker tl;dv leaked 181K meetings, sat on the fix for six months
Security researcher bobdahacker publicly disclosed that tl;dv, a popular AI notetaker for Zoom, Google Meet, and Teams, left 181,874 meeting records across 84,312 users and 35,003 email domains queryable by any authenticated user due to a missing Firestore tenant-isolation rule. Roughly 1,000 records were public and 715 invitee emails exposed; the researcher reported the flaw January 28, 2026 and it remained unfixed through repeated follow-ups, with active recording sessions including government-agency, university, HubSpot, and Confluent calls joinable by outsiders.
Corma raises $60M to build a defensive-only AI security lab
Corma, a Tel Aviv and San Francisco startup positioning itself as the first frontier lab for defensive cybersecurity AI, announced a $60M seed led by Sequoia Capital with Khosla Ventures and Coatue. CEO Alon Pluda says the company's models specialize in log and audit analysis rather than coding, and early Fortune 100/500 deployments cut threat response times by 94%.
AI agents finish whole online-course quizzes, chatbots don't refuse
The New York Times reports AI agents are moving online-course cheating beyond chatbot-written essays into fully automated coursework, executing prompts like 'log in and complete my quiz' end-to-end. Major chatbots reportedly did not refuse the prompts, raising fresh questions about the integrity of online degrees as agentic browsing becomes standard consumer capability.
iPhone 18 Pro parts cost jumps 38% as memory eats a third
TrendForce estimates the 256GB iPhone 18 Pro's bill of materials will run about 38% higher than the iPhone 17 Pro, with memory rising from roughly 10% of BOM a year ago to 34% in Q3 2026 and topping 40% by H1 2027. Memory prices have surged five- to sevenfold since early 2025 as AI infrastructure buyers absorb global HBM and DRAM output, forcing Apple to sacrifice gross margin instead of fully passing costs to consumers.
Meta releases Muse Glimmer, a 30B agent model that runs on a laptop
Meta released Muse Glimmer, a 30B-parameter dense multimodal model under Apache 2.0, tuned for local agentic tool use, coding, and LLM-as-judge with a 131K context and support for 100+ languages. 4-bit quantization compresses it under 20GB so it runs on a single consumer GPU, hitting 3.1x speedup on RTX 5090 via speculative decoding. Meta paired the drop with a 6,500-word Zuckerberg essay promising open weights for Muse Spark 1.2 in the coming weeks, defending model distillation, and a $1B community fund for regions hosting Meta data centers.
AI research system finds novel HTTP desync attacks in 700 live sites
PortSwigger's James Kettle unveiled HTTP Terminator, an autonomous AI research system that tested 30,000 candidate desync vectors against thousands of authorized websites and identified roughly 700 vulnerable targets, including banks, government infrastructure, security products and an airport. The system generated new attack classes including a dual-matching Content-Length pattern, a 'dangling-byte' technique for more reliable response queue poisoning, and shared-parser confusion, and a human-guided cascade also exposed an Apache Traffic Server zero-day. Kettle frames it as the first case of AI producing genuinely novel security research, rather than reapplying known bug classes.
Cloudflare's Kitesurf agent browser uses up to 7x less memory than Chromium
Cloudflare launched Kitesurf, a cloud-hosted browser purpose-built for AI agents that runs inside Workers V8 isolates. Built in 12 weeks by stitching together Blitz (renderer), Firefox's Stylo (CSS), Parley (text) and Boa JS, Kitesurf passes ~215,000 Web Platform Tests and reports 3.1x-3.8x less CPU and 4.7x-7.0x less memory vs Chromium for screenshotting and HTML extraction. Available free in beta via Browser Run; the pitch is that agents don't need themes, tabs or extensions and would rather trade rendering fidelity for token-cost and context-window efficiency.
Rippling built an AI cost tracker after AI spend hit 40% of R&D budget
Rippling launched AI Spend Console after its own AI-token bill was on track to consume 40% of R&D headcount budget, growing 80% month-over-month with 10-15% of employees driving 60% of spend and one engineer burning $50K/month. The tool maps spend per employee and team against productivity signals (code output, PRs) and routes across Cursor, OpenAI, Anthropic, Grok and Z.ai's GLM 5.2 — which CEO Parker Conrad calls '85% cheaper but nearly identical performance.' Token spend dropped from 40% to 15% of headcount budget; July costs were 37% of April despite similar 600B token volumes.
Anthropic loosens Fable 5 on biology, cuts blocked queries by 85%
Anthropic rewrote and retrained Fable 5's biology safety classifier to distinguish everyday health, education and clinical questions from dual-use research. The company says the change cuts biology-related fallbacks by about 85% and total fallback volume by ~67% on Claude.ai, 55% on Cowork, 17% on Claude Code and 7% on the Claude Platform. Virology, toxicology and molecular-design prompts still route to Opus 5, so Anthropic warns Fable 5 remains 'not yet usable for professional biology research and drug development.'
Claude Code makes auto mode the default, and human reviewers look worse for it
Anthropic will flip Claude Code's Auto Mode on by default for Pro, Max and Team users starting August 14, replacing manual approval prompts with a classifier that vets each tool call for irreversible or destructive actions. Anthropic says internal testing across 1,000+ paid users showed the classifier caught 89% of dangerous commands compared to just 13.6% for human reviewers, and teams using auto mode ship roughly 25% more pull requests. The company will stop charging for the extra tokens the classifier consumes.
Anthropic lets Claude Code sessions message each other
Claude Code v2.1.224 introduces cross-session messaging: one session can now send a summary to another mid-task rather than forcing users to re-explain context. Claude composes the actual message from a user hint, so it's coordination rather than a raw history dump. Permission approvals and configuration changes are excluded, and any privileged actions still prompt the receiving session. macOS and Linux only for now.
Stanford's AI designs 16 novel bacteria-killing viruses
Stanford researchers published in Science on Aug 6 the design of 16 novel bacteriophages generated by their genomic AI model Evo, which was trained on ~2M viral genomes. The AI-designed viruses successfully infected E. coli — some killing bacteria faster than the natural ΦX174 template. The team excluded human-pathogen data, but biosecurity experts including Dr. Moritz Hanke warn the technique could lower barriers for bioweapon design; Imperial's Tom Ellis says built-in genetic-data restrictions could mitigate the risk.
OpenAI's first gadget is a $300+ doughnut speaker, Gurman says
Bloomberg's Mark Gurman reports OpenAI's first consumer device — designed with Jony Ive's LoveFrom — is a battery-powered, screenless smart speaker roughly the size of a hockey puck, doughnut-shaped, with a camera, microphones and mechanical parts that shift so it appears 'alive.' Pricing is pegged at $300-$400 with a targeted 2027 ship, though Apple's trade-secrets suit over metal finishing techniques could delay it. The pitch: a portable, humanlike ChatGPT companion for the home.
AMD buys Taalas, which etches AI models into silicon
AMD said Wednesday it acquired Taalas, a Toronto startup founded in 2023 that bakes AI model weights directly into custom silicon rather than storing them in HBM. Terms were not disclosed; the deal is expected to close in Q4 2026 subject to regulatory approval. Taalas's first test chip HC1, on TSMC's 6nm process, hit ~17,000 tokens/sec serving Llama 3.1 8B — a claim of roughly 48x Nvidia GPUs and 8.5x Cerebras at time of announcement — and a 20B-parameter HC2 is due this summer.
Mathematicians say OpenAI's proofs plagiarize prior work
Steven Miller (Yeshiva) says the sphere-packing proof OpenAI showcased pastes in the central argument from his own 2016 paper without credit, calling the pattern deliberate. Francesco Fournier-Facio (Cambridge) says the soficity 'breakthrough' stitches together ideas from 2016 and 2019 papers. Follows the July 30 OpenAI drop of 10 Astra proofs previously covered as a breakthrough.
OpenAI hands free users unlimited chat and a Think button
OpenAI is making GPT-5.6 Luna the default for Free and Go users this week, replacing GPT-5.5, and rolling out unlimited text chats next week along with a new Think button for higher reasoning. OpenAI's internal evals put Luna's factual-error rate 62% below GPT-5.5-Instant and Sol's 68% below. File, image, voice and image-generation limits stay in place.

This Week's Biggest Movers in AI

vs average of last 4 issues — click to explore

Entity Now Avg Change
OpenAI 70 4.5 ▲ +1456%
Agents 70 5.8 ▲ +1107%
Chips 57 0.5 ▲ +11300%
Funding 49 2.3 ▲ +2030%
Regulation 49 3.8 ▲ +1189%
AI Infrastructure 47 2 ▲ +2250%
Anthropic 47 4.8 ▲ +879%
Google 37 1.8 ▲ +1956%
NVIDIA 29 1 ▲ +2800%
Generative AI 30 1.3 ▲ +2208%
Coding Tools 18 0 ▲ +100%
Inference 17 0.3 ▲ +5567%
Cybersecurity 17 0.8 ▲ +2025%
Microsoft 15 0.3 ▲ +4900%
Open Source 16 0.8 ▲ +1900%

How AI News Coverage Shifted This Week

News mix this week vs last

AI News Volume by Quarter

4-Issue Trend Lines: 113 AI Entities

Last 4 issues — click to explore