Or OpenAI's blog https://t.co/rnSrr68wy0 And attend today's Jalapeño Hot Chips session with Richard, Ravi, and Chris from OpenAI too. (7/7)
SemiAnalysis
Tracked through public AI activity and peer connections inside the directory.
- AI signals
- 11 past 30d
- Sources
- 5 distinct domains
- Discussions
- 0 past 30d
- Latest signal
- 4d ago
Articles & links
Read more at our article👇️ (6/7) https://t.co/dbuAUrRSt1
Gemini is Cooked but GCP is Cooking GCP YoY rev growth >100%, DeepMind's long term failure is Google Cloud's short term gain https://t.co/PqVUCF8qSg
Resources mentioned: 🟠 Using group theory to explore the space of positional encodings for attention (blog): https://t.co/ciWX1bLAWu 🟠 Positional Encodings and Group Theory (video): https://t.co/16BvKWPBU2 🟠 RoFormer: https://t.co/ZXMDVyvCf4 (7/7)
Are Open Models Catching Up? Comparing open vs. closed models across the eras of frontier models, Is the gap narrowing? https://t.co/TBZnvquwio
Everyone Says Datacenter Moratoriums Are Killing the US Buildout. We disagree. 300+ moratoriums mapped, 20GW sits inside a restricted local boundary, 1,525MW actually slips, 2.3GW nationwide including New York https://t.co/lYwyxj3XE7
- SemiAnalysis estimates only 2.3 GW of US datacenter capacity is genuinely delayed by moratoriums, against a projected 38 GW of new IT capacity in 2027.
- Of 20 GW planned inside restricted boundaries, just 1,525 MW is directly blocked at the local level, or 7.6%; New York's permit halt accounts for most of the rest.
- Michigan leads with 45 enacted moratoriums and Ohio has 40 (83% adopted in 2026), even as voters view datacenters unfavorably by 46% to 29%.
AgentX - InferenceXv3: Does CUDA Moat Hold up in Agentic Inferencing? $3 Million USD dataset open sourced, 1 Mil+ Context Length, Multiturn, Sub Agents 95%+ KVCache HitRate, GB300 NVL72, MI355, B200 https://t.co/PgtKNYXdFr
- SemiAnalysis's AgentX 1.0, built on 393 anonymized Claude Code traces at 1M+ context, cost more than $3M and used ~2MW across 1000+ chips.
- On Qwen3.5 SGLang the report puts Nvidia at 'over 20x better performance' at 90 tok/s/user; B300 FP4 shows '12x better performance per dollar' vs H100.
- AMD's ATOM stack shows strong single-GPU kernels but almost no production adoption — only one Alibaba ad unit runs it live, the authors say.
TokenBudgeting: Our Conversations with Enterprises on Token Spend Was Widespread TokenMaxxing Ever Really Here? https://t.co/tgNssBBGVk
- Meta employees consumed over 60 trillion tokens in a 30-day window in early 2026, with one individual alone accounting for about 280 billion.
- Monthly per-employee caps now range from $250 at an aerospace and defense manufacturer to $2,000 at Workday and Stripe, with no cross-industry consensus.
- Ramp data cited by SemiAnalysis shows 99th percentile customers spend about $90,000 per employee per year while the median customer spends $136.
Full Rubin results👇️ (2/2) https://t.co/Pkuj51yIOs
A Brain Too Big to Carry — On-Device vs Datacenter Inference Robot Models, Silicon & DRAM Efficiency, Jetson Thor vs. B300 TCO, Deployments, The Network Wall https://t.co/FzaOQl22RP
- Boston Dynamics offloads its System 2 planner to Google TPUs; the model runs at hundreds of billions to a trillion parameters, too big for a robot.
- For 96 robots, aggregate TCO is $14.97/hr on-device with Jetson Thor, $15.61 on RTX 6000 Pro offload, and $18.63 on B300.
- Sunday Robotics pivoted to on-device inference after home WiFi jitter proved unreliable; its ACT-2 model now reports 99.1% laundry-folding success.
To learn more, please check our networking model (7/7) https://t.co/kqt0IiSiFi
- SemiAnalysis is offering device-level tracking of AI cluster networking across five fabric layers, with data running 2023 to 2026.
- Coverage spans 80+ hyperscaler configuration panels for Microsoft, Google, Meta, Amazon, Oracle, X.AI and neoclouds, tied to specific accelerator SKUs.
- 25+ suppliers are tracked, including Nvidia, Arista, Broadcom, Cisco, Coherent and Lumentum, across 200G to 1.6T transceiver speeds.
Nvidia’s Backstop Universe Heads I Win, Tails Who Loses? The $11T AI Buildout, Nvidia’s Backstop Economics, and the Limits of Nvidia’s Balance Sheet https://t.co/cpGvPR0sUf
- Nvidia's 2Q F1/27 10-Q discloses $530B in gross off-balance-sheet guarantees, up from $184B the prior quarter.
- A single $108.5B line covers SB Energy's PORTS-Pike campus in Ohio: 4.25 GW leased to OpenAI for twenty years.
- AICP take-or-pay floors are disclosed for Firmus ($21.1B), SharonAI ($4.2B) and an estimated $2.2B for GMI.
Are you SemiAnalysis? Show it.
Add the Who’s Who of AI badge to your site or bio. It links back to this profile.
Markdown: [](https://aiweekly.co/whos-who/person/semianalysis)