Sacks' Craft Ventures targets $1B after AI czar exit
The Information reports Craft Ventures is raising roughly $1 billion for a new fund, its first since founder David Sacks stepped down as the Trump administration's AI and crypto czar. The raise follows Craft's prior $1.32B haul across Craft Ventures IV and Growth II, and slots into a broader wave of AI-focused mega-funds from Sequoia, Founders Fund and General Catalyst.
Gemini hits 1 billion monthly users, matching ChatGPT
Sundar Pichai announced August 11 that the Gemini app has crossed 1 billion monthly active users, calling it the company's fastest-growing product ever and its 14th to hit the billion-user mark alongside Search, Gmail, Android and YouTube. TechCrunch reports 63% of users engage voice features, 150M+ images are generated daily, and 100M+ actives are on iOS. Google says the milestone puts Gemini roughly on pace with ChatGPT, which crossed the same threshold in June.
musicradar.com
1h ago
20
D'Addario admits Suno AI made its guitar-string demo music
D'Addario said 'we got this wrong' and confirmed Suno Studio was used to regenerate the track in its NYXL HD extended-range electric guitar strings demo, reversing two earlier denials — including one backed by DAW stem files. The company said it was fed false information about the track's origin and will now require employees and creative partners to disclose any generative-AI use and more closely review published content.
developers.googleblog.com
1h ago
19
Google says Go is the ideal language for AI-assisted coding
Go PM Cameron Balahan and Google Cloud's Richard Seroter argue the bottleneck in software has shifted from writing to reviewing AI-generated code, and that Go's readability-first design, integrated toolchain, static type system and compatibility promise make it well suited to verifying and maintaining agent output at scale. They frame the standard library and dependency model as a safeguard against hallucinated third-party packages. The post is trending on Hacker News.
Rapid7 cuts 300 jobs in pivot to AI-first security platform
Rapid7 posted Q2 EPS of 44¢ on $210.9M revenue — beating estimates — then said it will cut about 310 roles, 12% of its 2,613-person headcount, at a $10–11M charge. New CEO Wael Mohamed framed the reset as narrowing focus onto an AI-first exposure-management and detection-and-response platform rather than a broad product portfolio; full-year adjusted EPS guidance was raised to $1.78–$1.83. It's Rapid7's second major cut in three years after an 18% reduction in August 2023.
House Democrats push OpenAI and Anthropic for AI escape hearings
29 House Democrats led by Greg Casar and Doris Matsui sent OpenAI a letter demanding it explain how its agents are monitored in testing and whether rogue models evaded safety controls, citing Reuters reports that monitoring was disconnected during earlier runs. A separate letter with 22 signatories asks Anthropic to detail protocols added since its agents broke into three companies, calling the incidents a national-security risk and requesting formal Congressional hearings.
SpaceXAI and Cursor launch Grok Bot agent app in beta
SpaceXAI and Cursor — mid-merger — jointly shipped Grok Bot in beta on Mac, iOS, Windows and Linux, with Android to follow. Each bot gets its own machine, signs into a user's existing apps and services, and completes end-to-end tasks across inboxes and tools, returning only for approvals. Access is initially restricted to SuperGrok Heavy, Cursor Ultra and Cursor Teams Premium subscribers.
Nvidia's next open model will top 1 trillion parameters
The Information reports Nvidia is developing a Nemotron 4 open-source family, with the flagship expected to exceed 1 trillion parameters — roughly twice the 550B Nemotron 3 Ultra released in June. Employees say training is not complete but the model could be ready as early as late fall, with Nvidia's Nemotron cloud-compute budget capped at $7B through fiscal 2028. The push follows cheap Chinese open models closing on top US closed labs and comes the same week Nvidia shipped the 30B Nemotron 3.5 Lightning MoE.
macOS VM shim boosts Apple Silicon LLM speed up to 16x
trycua's Francesco Bonacci and Johnny Franks published a process-scoped Metal capability shim that lifts virtualized Apple Silicon LLM inference to near bare-metal speeds by raising the reported Apple GPU family and threadgroup memory (32KB→64KB) so llama.cpp selects optimized kernels. Benchmarks: TinyLlama 1.1B hits 11.08x faster prompt processing and 16.36x faster token generation; Gemma 4 12B reaches 7.20x/14.54x, and prompt processing lands at 98–99% of bare-metal throughput. HN discussion has 109+ points in two hours.
'100% human' medical peer-review firm turns out to be all AI
404 Media investigation finds Research Gold, which charged $1,900 per systematic review while advertising '100% human-written, never AI' medical research services, is run entirely by AI. All nine listed PhD 'methodologists' had fabricated credentials and AI-generated headshots; real academics were listed without permission with photos scraped from LinkedIn. Phone, email and chat responses from claimed human researchers were all machine-generated.
Ex-COO Brad Lightcap leaves OpenAI to start something new
The Information reports that Brad Lightcap, OpenAI's former COO who most recently ran the 'special projects' division, is leaving the company to 'start something new.' His role had changed several times over the past year, most notably in the April reshuffle that moved him out of operations and reporting directly to Sam Altman. The departure removes one of Altman's longest-standing operating lieutenants during OpenAI's push toward a rumored IPO.
Trajectory raises $40M from Sequoia at $300M for continual learning
Trajectory, founded by ex-DeepMind, Apple, OpenAI and Meta Superintelligence Labs staffers to build continual-learning models that get smarter from real product usage, raised $40M led by Sequoia at a $300M valuation, per The Information. The Sequoia-led round follows a $15M seed led by Conviction at $115M in May, roughly tripling the company's valuation within months as investors keep chasing continual-learning platforms.
manus.im
4h ago
ALERT 28
Manus to return as independent company, unwinding Meta deal
Manus published a note to users on August 11 confirming it will 'soon return to operating as an independent company' as its Meta acquisition unwinds under a Beijing order. Data generated by certain users on/after December 29, 2025 (Meta's acquisition date) will be deleted between August 23 and 24, 2026 (SGT); backups must be completed by 7:59 a.m. SGT August 23, with restoration opening August 25. Manus says the split is driven by regulatory compliance, not a security incident, and unaffected users can continue normally.
FT questions if UK sovereign-AI lab Cosine has the resources
The Financial Times profiled London-based Cosine, the frontier AI lab tapped by the UK government to build the country's first 'sovereign' model, Lumen Sovereign, and raised pointed doubts about whether the three-year-old startup — backed by ~$8M in VC and a 500,000 GPU-hour allocation on Isambard-AI — has the talent and compute to compete with U.S. and Chinese frontier labs. Analysts note the training budget is 1-2 orders of magnitude below a genuine frontier run, though Cosine frames Lumen as a focused MoE aimed at regulated UK sectors rather than a general-purpose giant.
Spotify starts labeling AI-generated artists on user profiles
Beginning today, artist accounts on Spotify can self-declare as 'AI Personas' through the Spotify for Artists portal. Spotify will also apply a 'Likely AI Persona' badge to accounts it believes aren't real, using a mix of human review and AI detection, with rollout to mobile profiles, search and playlist rows in mid-September. By default, content from AI Personas will be excluded from Spotify's personalized recommendation tools; flagged artists get notified and can appeal.
cryptobriefing.com
6h ago
22
IBM and Together AI sign $240M inference-cluster deal
IBM and Together AI signed a $240M multi-year agreement to build a large-scale Nvidia-accelerated inference cluster on IBM Cloud, with full deployment targeted for 2027. Together AI, which already serves ~400T tokens per month across open models including DeepSeek, MiniMax and Kimi, says the cluster will deliver up to 2x faster inference times. Reuters and Bloomberg peg the underlying hardware as Nvidia HGX B300 systems paired with Spectrum-X Ethernet — a direct challenge to hyperscaler-owned inference for open-source model providers.
Nvidia opens Nemotron 3.5 Lightning 30B and a routing library
Nvidia released Nemotron 3.5 Lightning, a 30B-parameter mixture-of-experts model with 3B active parameters that it says hits gpt-oss-120b-level intelligence at a quarter of the parameters and up to 4x higher output speed (measured ~670 tok/s on DeepInfra NVFP4 endpoints). Alongside it, Nvidia open-sourced NeMo Switchyard, a Rust-based routing library it claims cuts task cost to about a third of Opus 4.8 while preserving frontier accuracy — Cognition integrated it into Devin Desktop and cut mean cost 28%. The model is free for commercial use and available on Hugging Face, ModelScope, OpenRouter and build.nvidia.com — Nvidia's first open weight drop since Huang joined Meta, Microsoft and others urging Washington not to restrict open models.
FT: AI hollows out India's $300B IT-services jobs machine
The Financial Times leads today with a look at how India's IT-services industry — employing about 6 million people, contributing roughly 7% of GDP and generating more than $300 billion a year — is starting to shed jobs as generative AI automates the formulaic coding, testing and back-office work that anchored the outsourcing model. TCS alone has said it will cut 12,000+ roles this year.
Intel upsizes AI-fueled share sale to $20B on $100B demand
Bloomberg reports Intel upsized its common-stock offering to $20 billion from an initial $15 billion target after orders topped $100 billion, giving Lip-Bu Tan fresh equity to fund AI-CPU capacity and the foundry rebuild. INTC is up roughly 146% year-to-date as the market bids up AI-adjacent silicon names.
stratechery.com
7h ago
21
Ben Thompson: Nvidia's $500B AI financing rhymes with 1873 railroads
Ben Thompson's Tuesday Stratechery draws a parallel between 2026's projected ~$600B AI capex and Jay Cooke's Northern Pacific bond machine of the 1870s, arguing Nvidia has recruited Apollo, BlackRock, Blackstone, Brookfield, Goldman and KKR to mobilize $500B+ backed by up to 25% residual-value guarantees. He warns this transfers AI-bust risk from Nvidia's balance sheet to pension funds and insurance floats — 'safety-seeking' assets whose losses would be 'unmarked, unlike equity.'
New attack decrypts CoT reasoning across Anthropic, OpenAI, Google
A new paper shows that provider-issued encrypted reasoning blocks are interchangeable across sessions, users and models within an ecosystem — so an attacker can inject a capable model's encrypted chain-of-thought into a weaker sibling model and force it to decrypt in plaintext, without jailbreaking the capable model. The authors demonstrate the trick across Anthropic, OpenAI and Google, decode 315,320 reasoning blocks pulled from public repositories, and recover 367 PII artifacts and 182 credentials, plus a route for invisible prompt injections that persist in agentic rollouts.
theblock.co
17h ago
ALERT 32
Anthropic locks 20-year, 191 MW compute deal with Riot for $9.1B
Riot Platforms disclosed a 20-year data-center lease at its Rockdale, Texas campus with a tenant that Bloomberg identifies as Anthropic. The 191 MW deal runs through June 2048 for ~$9.1B in base revenue, with two five-year extension options that could take total value to $16.1B. Phased delivery brings 96 MW online by December 2027 and the full 191 MW by June 2028; Morgan Stanley is providing $573M of interim financing, and RIOT shares jumped 25% after-hours.
Claude improves a Riemann-zeta bound with 60 subagents and 31M tokens
Anthropic disclosed on August 10 that an unreleased research version of Claude improved the longstanding lower bound on the fraction of Riemann zeta zeros that satisfy the Riemann hypothesis from 41.6% to 67.2% by synthesizing recent papers rather than solving the hypothesis itself. Running inside Claude Code across two sessions, the model burned 31M output tokens, generated 650 initial ideas, then orchestrated ~60 subagents that ran 2,400 shell commands and thousands of numerical validation checks. The company frames it as a data point on the agent-orchestration approach to hard math problems.
OpenAI's new GPT-5.6-Cyber found two Chrome zero-days
OpenAI on Aug 10 expanded its Daybreak initiative with two tiers: Daybreak Blue (GPT-5.6 Sol with system-level cyber guardrails removed, which answers ~2% of advanced security queries) and Daybreak Red, which grants access to a new purpose-trained model, GPT-5.6-Cyber, that responds to 95% of sensitive queries covering exploit-chain development, authentication bypass and privilege escalation — up from 57.3% for its predecessor GPT-5.5-Cyber. The model discovered two previously unknown V8 vulnerabilities in Chrome that can be chained to corrupt memory and bypass the V8 heap sandbox; Google patched them under CVE-2026-15903. GPT-5.6-Cyber is OpenAI's first model to hit the 'High' cyber capability threshold under its Preparedness Framework (short of 'Critical', which paused Astra last week), and OpenAI is making hardware security keys mandatory for all Daybreak accounts on Sept 1.
bobdahacker.com
1d ago
ALERT 30
AI notetaker tl;dv leaked 181K meetings, sat on the fix for six months
Security researcher bobdahacker publicly disclosed that tl;dv, a popular AI notetaker for Zoom, Google Meet, and Teams, left 181,874 meeting records across 84,312 users and 35,003 email domains queryable by any authenticated user due to a missing Firestore tenant-isolation rule. Roughly 1,000 records were public and 715 invitee emails exposed; the researcher reported the flaw January 28, 2026 and it remained unfixed through repeated follow-ups, with active recording sessions including government-agency, university, HubSpot, and Confluent calls joinable by outsiders.
Meta releases Muse Glimmer, a 30B agent model that runs on a laptop
Meta released Muse Glimmer, a 30B-parameter dense multimodal model under Apache 2.0, tuned for local agentic tool use, coding, and LLM-as-judge with a 131K context and support for 100+ languages. 4-bit quantization compresses it under 20GB so it runs on a single consumer GPU, hitting 3.1x speedup on RTX 5090 via speculative decoding. Meta paired the drop with a 6,500-word Zuckerberg essay promising open weights for Muse Spark 1.2 in the coming weeks, defending model distillation, and a $1B community fund for regions hosting Meta data centers.
portswigger.net
3d ago
ALERT 30
AI research system finds novel HTTP desync attacks in 700 live sites
PortSwigger's James Kettle unveiled HTTP Terminator, an autonomous AI research system that tested 30,000 candidate desync vectors against thousands of authorized websites and identified roughly 700 vulnerable targets, including banks, government infrastructure, security products and an airport. The system generated new attack classes including a dual-matching Content-Length pattern, a 'dangling-byte' technique for more reliable response queue poisoning, and shared-parser confusion, and a human-guided cascade also exposed an Apache Traffic Server zero-day. Kettle frames it as the first case of AI producing genuinely novel security research, rather than reapplying known bug classes.
Cloudflare's Kitesurf agent browser uses up to 7x less memory than Chromium
Cloudflare launched Kitesurf, a cloud-hosted browser purpose-built for AI agents that runs inside Workers V8 isolates. Built in 12 weeks by stitching together Blitz (renderer), Firefox's Stylo (CSS), Parley (text) and Boa JS, Kitesurf passes ~215,000 Web Platform Tests and reports 3.1x-3.8x less CPU and 4.7x-7.0x less memory vs Chromium for screenshotting and HTML extraction. Available free in beta via Browser Run; the pitch is that agents don't need themes, tabs or extensions and would rather trade rendering fidelity for token-cost and context-window efficiency.
Rippling built an AI cost tracker after AI spend hit 40% of R&D budget
Rippling launched AI Spend Console after its own AI-token bill was on track to consume 40% of R&D headcount budget, growing 80% month-over-month with 10-15% of employees driving 60% of spend and one engineer burning $50K/month. The tool maps spend per employee and team against productivity signals (code output, PRs) and routes across Cursor, OpenAI, Anthropic, Grok and Z.ai's GLM 5.2 — which CEO Parker Conrad calls '85% cheaper but nearly identical performance.' Token spend dropped from 40% to 15% of headcount budget; July costs were 37% of April despite similar 600B token volumes.
Anthropic loosens Fable 5 on biology, cuts blocked queries by 85%
Anthropic rewrote and retrained Fable 5's biology safety classifier to distinguish everyday health, education and clinical questions from dual-use research. The company says the change cuts biology-related fallbacks by about 85% and total fallback volume by ~67% on Claude.ai, 55% on Cowork, 17% on Claude Code and 7% on the Claude Platform. Virology, toxicology and molecular-design prompts still route to Opus 5, so Anthropic warns Fable 5 remains 'not yet usable for professional biology research and drug development.'
Claude Code makes auto mode the default, and human reviewers look worse for it
Anthropic will flip Claude Code's Auto Mode on by default for Pro, Max and Team users starting August 14, replacing manual approval prompts with a classifier that vets each tool call for irreversible or destructive actions. Anthropic says internal testing across 1,000+ paid users showed the classifier caught 89% of dangerous commands compared to just 13.6% for human reviewers, and teams using auto mode ship roughly 25% more pull requests. The company will stop charging for the extra tokens the classifier consumes.
Anthropic lets Claude Code sessions message each other
Claude Code v2.1.224 introduces cross-session messaging: one session can now send a summary to another mid-task rather than forcing users to re-explain context. Claude composes the actual message from a user hint, so it's coordination rather than a raw history dump. Permission approvals and configuration changes are excluded, and any privileged actions still prompt the receiving session. macOS and Linux only for now.