Home › AI Use-Case Library › AI in Software & Tech: 155 real deployments

AI in Software & Tech: 155 real deployments

Named Software & Tech organisations and what they run, grouped by business function.

155deployments
111in production or with results
79with a reported outcome
12halted or reversed
Sep 21, 2026last updated

Software development 25 deployments

Clay

uses Raindrop Simulations to replay production AI agent traces against proposed code changes using synthetic copies of databases, payment APIs, and communications tools to catch hallucinations, tool misuse, and behavior drift in pull requests

In production Sep 18, 2026 Simulations Source: runtimewire.com
Framer

uses Raindrop Simulations to replay production AI agent traces against proposed code changes using synthetic copies of databases, payment APIs, and communications tools to catch hallucinations, tool misuse, and behavior drift in pull requests

In production Sep 18, 2026 Simulations Source: runtimewire.com
Vercel

uses Raindrop Simulations to replay production AI agent traces against proposed code changes using synthetic copies of databases, payment APIs, and communications tools to catch hallucinations, tool misuse, and behavior drift in pull requests

In production Sep 18, 2026 Simulations Source: runtimewire.com
Z.ai

Deployed a GLM-5.3 Infra Agent to build and optimize the inference infrastructure for GLM-5.3-Flash on a cluster of more than 100,000 Chinese-made AI accelerators

Reported: End-to-end throughput tripled from baseline in under two weeks; per-token cost and hardware efficiency described as comparable to mainstream Nvidia GPUs

Results reported Sep 17, 2026 GLM-5.3 Source: z.ai
Google

Google opened Claude Opus 5 access to all engineers through its internal Antigravity IDE for software development, replacing a prior policy requiring Gemini use

In production Sep 14, 2026 Claude Opus 5, Antigravity Source: businessinsider.com
NVIDIA

using Temporal Durable Execution platform for AI agent workloads

In production Sep 14, 2026 Source: temporal.io
OpenAI

using Temporal Durable Execution platform to keep long-running AI agents fault-tolerant

Reported: OpenAI's Temporal usage grew 60-fold in under a year

Results reported Sep 14, 2026 Source: temporal.io
GitHub

Launched Project HydraFusion multi-model orchestration system routing coding tasks across multiple LLMs using direct-solve, cascade, and critique-and-revise patterns, available as research preview in Copilot CLI

Reported: 4.9 pt gains on TerminalBench 2.1 at 67% lower cost vs Claude Opus 5; 36% cost reduction on DeepSWE; 65% lower cost on CheckpointBench

Pilot Sep 4, 2026 HydraFusion Source: github.blog
Meta

Graduated Muse Code terminal AI coding agent from beta to production, adding inter-session messaging over Unix sockets, multi-subagent workflow orchestration, and a TypeScript SDK

In production Aug 31, 2026 Muse Code Source: developer.meta.com
DoltHub

Used AI agents to author approximately 2,000 pull requests to build DoltLite Beta, a SQLite fork with Git-style versioning

Reported: DoltLite Beta shipped, built via roughly 2,000 AI-agent-authored pull requests

Results reported Aug 31, 2026 Source: dolthub.com
Grindr

AI writes approximately 70% of Grindr's code across engineering operations

Reported: Engineering output has 2.5x'd since July 2025

Results reported Aug 29, 2026 Coding agents Source: ft.com
Cursor

AI coding assistant using Anthropic, Google, and SpaceXAI models in production; OpenAI model supply being wound down effective November 12 following SpaceX acquisition

Reported: OpenAI represents only ~5% of Cursor's traffic

In production Aug 28, 2026 Coding agents Source: cybersecuritynews.com
ByteDance

Consolidating Trae coding platform and Coze agent-building tool into Doubao super-app, and planning to launch Doubao Work productivity agent

Announced Aug 24, 2026 Trae, Coze, Doubao Coding agents Source: bloomberg.com
Warp

Uses Factories, version-controlled pipelines that move tickets through spec, implementation, review and verification with coding agents, to handle internal tasks

Reported: factories already handle 30-35% of Warp's own internal tasks

Results reported Aug 18, 2026 Coding agents Source: warp.dev
Block

Built and uses Berd internally, a desktop app for managing AI agents, files, skills and sessions across the Goose framework

In production Aug 18, 2026 Goose Coding agents Source: cryptobriefing.com
Snowflake

Used GitHub Copilot Autofix to generate a security patch for snowflake-connector-net, which replaced a safe input pattern with raw string interpolation of a GitHub issue title

Reported: Autofix-generated patch introduced an exploitable shell-injection vulnerability; unauthenticated attacker exfiltrated Jira token for [email protected] within five days of the patch

Results reported Aug 17, 2026 GitHub Copilot Autofix Coding agents Source: wiz.io
GitHub

Integrated xAI's Grok 4.6 reasoning model into GitHub Copilot for agentic coding and multi-step workflows

In production Aug 14, 2026 Grok 4.6, GitHub Copilot Coding agents Source: github.blog
Cognition

Integrated Nvidia NeMo Switchyard router into Devin Desktop to reduce AI inference cost

Reported: cut mean cost 28%

Results reported Aug 11, 2026 NeMo Switchyard Coding agents Source: blogs.nvidia.com
Rippling

deployed AI Spend Console to map per-employee and per-team AI token usage against productivity signals and route spend across multiple AI providers to reduce costs

Reported: token spend dropped from 40% to 15% of R&D headcount budget; July costs were 37% of April despite similar 600B token volumes

Results reported Aug 7, 2026 Cursor, Grok, GLM 5.2 Coding agents Source: techcrunch.com
Databricks

Databricks routes AI coding tasks by complexity, defaults away from frontier models when cheaper models clear the bar, and trims coding-harness prompt overhead to control enterprise coding-agent costs.

Reported: Databricks reports dynamic routing cut average task cost by more than 30% and harness tuning reduced generated tokens by almost 50%.

Results reported Aug 7, 2026 Coding agents Source: Databricks
1Password

1Password's engineering team used AI agents to autonomously refactor a large monolithic codebase, with human-oversight patterns for cross-file dependency tracking, test suite maintenance, and rollback logic.

Reported: Team reports meaningful velocity gains while flagging specific failure modes

In production May 15, 2026 Coding agents Source: 1Password Blog
EPAM Systems

EPAM is building a practice of 10,000 Claude-certified architects, including 250 forward-deployed engineer 'Black Belts', to deliver enterprise AI for Global 2000 clients using Claude models, Claude Code, and the Claude Agent SDK.

Reported: 1,300 architects already certified; 5,000 targeted by end of Q3 2026

In production May 7, 2026 Claude, Claude Code, Claude Agent SDK Coding agents Source: EPAM
LlamaIndex

LlamaIndex generates the overwhelming majority of its codebase with AI, per CEO Jerry Liu.

Reported: Roughly 95% of the company's codebase is now AI-generated

Results reported May 2, 2026 Coding agents Source: VentureBeat
Google

Google generates most of its new code with AI, with AI-written code human-reviewed, per Pichai's disclosure at Cloud Next.

Reported: AI-written, human-reviewed code went from ~30% (April 2025) to 50% (last fall) to 75% today, with one complex migration completed 6x faster year-over-year

Results reported Apr 22, 2026 Coding agents Source: Techmeme

Security operations 23 deployments

Hacktron AI

used Claude Opus 5 to chain two OpenAI vulnerabilities—a libheif memory bug reachable via HEIF image uploads on the Discourse-based OpenAI community forum—to take over multiple OpenAI employee accounts and access OpenAI's internal monorepo in authorized bug bounty research

Reported: successfully opened a pull request in OpenAI's internal monorepo; OpenAI paid a $6,500 bug bounty and fixed the flaws

Results reported Sep 18, 2026 Claude Opus 5 Source: techcrunch.com
Google

Ran Gemini in a capture-the-flag security evaluation conducted by Israeli firm Irregular; a misconfiguration allowed the model to reach the real internet

Reported: Gemini autonomously hacked three companies: brute-forced passwords at one, scraped credentials from public repos at two others; stopped hacking on its own in all three cases

Results reported Sep 18, 2026 Gemini Source: spokesman.com
Enclave

using DeepSeek V4.1 Flash as autonomous agent for offensive security testing against vulnerable targets

Reported: achieved code execution on all 11 vulnerable targets for $4.65 per run, executing 2,349 bash commands over 268.3M input tokens

Results reported Sep 16, 2026 DeepSeek V4.1 Flash Source: enclave.ai
ZRON

running AI systems over stolen foreign intelligence data to make it searchable and actionable for police clients

In production Sep 16, 2026 Source: wsj.com
Hugging Face

Hugging Face used GLM 5.2 to analyze the July 2026 security breach after commercial AI tools refused to process OpenAI-related analysis

In production Sep 15, 2026 GLM 5.2 Source: thenextweb.com
Strix

Strix used an AI agent to discover an unauthenticated Harbor registry and a live GITHUB_TOKEN with admin access to Baseten's repositories baked into a 2023 Docker image

Reported: AI agent found admin and push access to Baseten's main product repo, its GitOps deployment repo, and its Homebrew distribution channel

Results reported Sep 15, 2026 Source: strix.ai
AISLE

Autonomous vulnerability-finding system deployed against the curl production codebase

Reported: 29 reports produced; 6 accepted as CVEs: CVE-2026-80229, CVE-2026-80230, CVE-2026-80231, CVE-2026-80255, CVE-2026-82208, CVE-2026-82209

Results reported Sep 2, 2026 Source: aisle.com
CrowdStrike

Launched SafeMind dual-model agentic system where Red Tempest probes for attack paths and Blue Solano patches them in a closed loop on a digital twin of the customer environment, shipping inside Falcon with standalone access via Project QuiltWorks

In production Sep 1, 2026 SafeMind, Red Tempest, Blue Solano, Falcon Source: siliconangle.com
Anthropic

using Alice to red-team and monitor production AI systems for jailbreaks, prompt injection, and agentic misuse

In production Aug 25, 2026 Cybersecurity Source: techstartups.com
OpenAI

Implemented chain-of-thought monitoring on all RL training and evaluations involving tools for GPT-5.6 Sol-class and higher models and all inference on the Astra model

Reported: monitoring overhead at roughly 20% of the inference compute being monitored

Results reported Aug 19, 2026 AI safety operations Source: theregister.com
Wiz

Used Wiz Red Agent AI to audit Snowflake's snowflake-connector-net repository and discover a shell-injection vulnerability introduced by a GitHub Copilot Autofix patch

Reported: Red Agent identified an exploitable shell-injection flaw; unauthenticated attacker exfiltrated the Jira token for [email protected] within five days, granting read access to engineering, security compliance and bug bounty projects

Results reported Aug 17, 2026 Wiz Red Agent Cybersecurity Source: wiz.io
OpenAI

deployed GPT-5.6-Cyber through the Daybreak Red program for offensive security vulnerability research, discovering previously unknown Chrome V8 vulnerabilities

Reported: model discovered two previously unknown V8 vulnerabilities in Chrome that can be chained to corrupt memory and bypass the V8 heap sandbox, patched by Google under CVE-2026-15903; model responds to 95% of sensitive security queries versus 57.3% for predecessor GPT-5.5-Cyber

Results reported Aug 10, 2026 GPT-5.6-Cyber Cybersecurity Source: the-decoder.com
PortSwigger

deployed autonomous AI research system (HTTP Terminator) to test 30,000 candidate desync vectors against thousands of authorized websites to identify novel HTTP desync vulnerabilities

Reported: identified roughly 700 vulnerable targets including banks, government infrastructure, security products, and an airport; generated new attack classes including dual-matching Content-Length pattern, dangling-byte technique, and shared-parser confusion; exposed an Apache Traffic Server zero-day

Results reported Aug 7, 2026 Cybersecurity Source: portswigger.net
Google

Google uses AI tools to discover, validate, triage and fix Chrome vulnerabilities and is piloting twice-weekly security releases to absorb the resulting volume.

Reported: Chrome versions 149 and 150 fixed 1,072 bugs between them.

Results reported Jul 30, 2026 Cybersecurity Source: Wired
XBOW

XBOW used its autonomous offensive-security agent to test Microsoft Bing Images and disclose two unauthenticated command-injection flaws.

Reported: The agent found CVE-2026-32194 and CVE-2026-32191, both rated CVSS 9.8; Microsoft fixed both before disclosure.

Results reported Jul 24, 2026 XBOW autonomous security agent Cybersecurity Source: The Hacker News
Meta

Meta deployed a free Facebook identity badge that uses a facial-recognition selfie to distinguish real users from AI-generated impostors.

In production Jul 24, 2026 Facebook Verified Content moderation and AI detection Source: Meta
Searchlight Cyber

Searchlight Cyber used GPT-5.6 Sol Ultra for autonomous multi-agent analysis of the default WordPress codebase.

Reported: About 10 hours and $25 of model usage found a pre-authenticated SQL-injection-to-remote-code-execution chain affecting more than 500 million WordPress instances.

Results reported Jul 20, 2026 GPT-5.6 Sol Ultra Cybersecurity Source: Searchlight Cyber
OpenAI

OpenAI uses the internal GPT-Red model to automate red teaming and adversarially train frontier models against prompt injection.

Reported: Prompt-injection attacks fell from more than 90% success on GPT-5 to under 23% on GPT-5.6; GPT-Red reached 84% attack success versus 13% for human red-teamers on a benchmark.

Results reported Jul 15, 2026 GPT-Red Cybersecurity Source: MIT Technology Review
Microsoft

Microsoft is reorganizing its cybersecurity business around AI-assisted vulnerability discovery and automated threat response.

Reported: The restructuring replaced senior executives and cut hundreds of roles.

Results reported Jul 14, 2026 Security Copilot Cybersecurity Source: The Information
Microsoft

Microsoft uses the multi-model MDASH agentic scanning harness to discover Windows vulnerabilities faster.

In production Jul 10, 2026 MDASH Cybersecurity Source: The Register
Verkada

Physical-security company Verkada is adopting Nvidia's Cosmos world foundation models and Physical AI Data Factory toolkit to scale AI across its 2.4 million connected devices.

Reported: The integration has already delivered a 68% accuracy gain in spatial-temporal video search.

Results reported Jul 2, 2026 NVIDIA Cosmos, Physical AI Data Factory Surveillance and public safety Source: SiliconANGLE
Cloudflare

Under Project Glasswing, Cloudflare deployed Anthropic's Mythos model against real-world cyber frontier threat scenarios across Cloudflare infrastructure, publishing model behavior, capability boundaries, and detection findings.

Pilot May 18, 2026 Claude Mythos Cybersecurity Source: Cloudflare

Operations 20 deployments

Langfuse

Integrated TypeSafe's Jev decision model into its platform

In production Sep 19, 2026 Jev Source: forbes.com
LangChain

Integrated TypeSafe's Jev decision model into its AI workflow platform

In production Sep 19, 2026 Jev Source: forbes.com
Cloudflare

Integrated TypeSafe's Jev decision model into Workers AI, callable via plain REST POST reusing the Workers AI token

In production Sep 19, 2026 Jev Source: forbes.com
Vercel

Integrated TypeSafe's Jev decision model into AI Gateway for developer teams

Reported: reached nearly 13% of paid teams within 24 hours, roughly 2x the GPT-5.6 family and over 6x Fable 5.1 at the same adoption mark

Results reported Sep 19, 2026 Jev Source: forbes.com
Meta

Meta announced deployment of its MTIA 450 custom AI accelerator chips in data centers starting H1 2027, with MTIA 500 following by end of 2027, as part of a ~$115B capex plan

Announced Sep 15, 2026 MTIA 450 Source: bloomberg.com
Andon Labs

Ran persistent AI agents to autonomously operate vending machines, a retail store, and a café for nearly two years prior to public launch of Pion

In production Sep 14, 2026 Source: andonlabs.com
Palantir

limiting use of Anthropic Fable model following Anthropic's June policy change granting 30-day retention of usage logs

Halted / reversed Sep 14, 2026 Fable Source: theinformation.com
Nvidia

limiting use of Anthropic Fable model following Anthropic's June policy change granting 30-day retention of usage logs

Halted / reversed Sep 14, 2026 Fable Source: theinformation.com
Anthropic

Signed $45B compute deal with Nscale for AI cloud infrastructure

Announced Sep 4, 2026 Source: thenextweb.com
DeepSeek

Plans to deploy 160,000 Huawei Ascend 950DT accelerators at a 1GW data center in Inner Mongolia to serve AI inference workloads

Announced Sep 4, 2026 Source: bloomberg.com
xAI

Expanded Grok Bot to enterprise-wide deployment with admin controls, action recording, audit logging, and OpenTelemetry export for autonomous worker tasks

In production Sep 3, 2026 Grok Bot Source: superpowerdaily.com
Wafer

Deploys autonomous agents to profile and tune inference workloads across GPU hardware and model architectures

Reported: agent-tuned AMD MI355X GPUs hit approximately 80% of Nvidia B200 throughput

Results reported Sep 1, 2026 Source: cryptobriefing.com
Meta

Deployed real-time streaming ASR model processing audio in 80ms chunks across 25+ languages with diarization and endpointing for recordings over an hour with 20+ speakers, available via Meta Model API, Meta AI for Mac, and Muse Code

Reported: 3.1% word error rate on streaming, 17.5% diarization error rate

Results reported Sep 1, 2026 Muse Voice Transcribe Source: research.meta.ai
Perplexity

Deployed Hybrid Compute mode that auto-classifies prompts and routes sensitive content to on-device models while sending general work to cloud models, with no token fees for local runs

In production Sep 1, 2026 Gemma E4B, Qwen 3.6, Opus 5, GPT-5.6 Sol Source: engadget.com
Google Cloud

Deploying context-creating AI agents within its own tools to automate tasks traditionally handled by forward-deployed engineers embedded with enterprise customers

Google

Contributing AI contrail-prediction models and satellite-based verification to Operation Blue Skies, a 30-month trial guiding altitude adjustments for approximately 10,000 flights per year through Shanwick oceanic airspace across two four-month operational trials

SAP

SAP is deploying its Joule copilot and 100+ internal AI use cases while urging its ~108,000 employees to reinvent their roles with AI-augmented workflows rather than face outright layoffs.

Reported: CEO cites Joule copilot deployments and 100+ AI use cases as evidence the retraining bet is working; CFO disclosed annual cuts of 1-2% (~1,000-2,000 jobs) will now recur indefinitely.

In production Jul 2, 2026 Joule Enterprise copilots Source: The New York Times
OpenAI

OpenAI rolled out its Codex agent as the primary internal AI tool across every department, including Legal and Recruiting, shifting staff from single-prompt chat to multi-step agentic workflows.

Reported: 97.9% of employees now use Codex, up from ~40% in August 2025; non-developer usage surged 137x for individual users and 189x for organizational users; the median legal employee generated 13x more monthly output tokens in June 2026 than in November 2025.

Results reported Jun 26, 2026 Codex Enterprise copilots Source: The Register
Microsoft

Microsoft's internal program data covers its enterprise deployment of AI agents, where token consumption from agentic workflows burns through annual software budgets in months.

Reported: Internal data shows deploying AI agents costs more than employing human workers for equivalent tasks; enterprise AI ROI remains elusive

In production May 23, 2026 Operations automation Source: Fortune

Customer service 17 deployments

Apple

Testing Siri AI-centric smart home hub (J490) with facial recognition personalization in employees' homes ahead of consumer release

Pilot Sep 20, 2026 Siri Source: bloomberg.com
OpenAI

Rolling out ChatGPT for Teens to Australian users aged 13-17 with age-appropriate safeguards and parental controls

In production Sep 19, 2026 ChatGPT Source: openai.com
ByteDance

operating Doubao AI chatbot as a consumer product

Reported: 75M MAUs

Results reported Sep 16, 2026 Doubao Source: bloomberg.com
Anthropic

Leasing 2.16GW data center capacity in Queensland, Australia to serve Claude inference for user queries starting 2027

Announced Sep 16, 2026 Claude Source: abc.net.au
Google

Google deployed Gemini 3.8 Live and Extended Thinking for conversational agents in Search Live, Workspace, Gmail, and Keep

Reported: #1 on Artificial Analysis Speech-to-Speech Quality Index at 82.6, 68.6% on τ-Voice, 35.1% on Sierra's τ-Voice-banking, 97.7% on Big Bench Audio

Results reported Sep 15, 2026 Gemini 3.8 Live, Gemini 3.8 Live Extended Thinking Source: blog.google
Meta

Shipping closed-weight Muse Spark model across Facebook, Instagram, WhatsApp, and Meta AI

In production Sep 14, 2026 Muse Spark Source: colossus.com
ThunderPhone

testing GPT-Live-1 API for AI voice agents on a 13,000-token insurance-qualification script across roughly a dozen real phone calls

Reported: repeated instruction-following failures worse than earlier realtime models

Results reported Sep 13, 2026 GPT-Live-1 Source: reddit.com
Google

Gemini chatbot used by consumers for outdoor trip safety planning, providing Mount Shasta hikers with recommendations for food and water quantities

Reported: Three hikers required rescue after Gemini recommended far less food and water than their group required; Siskiyou County rangers urged hikers to consult local ranger stations rather than AI chatbots for safety-critical planning

Results reported Sep 5, 2026 Gemini Source: techcrunch.com
Google

Replaced Google Assistant with Gemini on Android phones, tablets, Wear OS smartwatches, and Android Auto for Hey Google triggers and power-button long presses

In production Sep 4, 2026 Gemini Source: implicator.ai
Google

Integrating Gemini into Android Find Hub with a 'Remembered' tab for location memory, Gemini Live Guided Vision for low-light and fine-print assistance, and Motion Assist to reduce motion sickness

Announced Sep 1, 2026 Gemini, Gemini Live Source: blog.google
Meta

Planning to launch Hatch, a paid consumer AI agent platform tiered up to $199/month, trained in a sandbox simulating DoorDash, Etsy, Reddit, Yelp, and Outlook

Announced Aug 24, 2026 Claude Opus 4.6, Sonnet 4.6 Consumer assistants Source: theinformation.com
Apple

Deploying a proprietary China-specific LLM trained with Alibaba support for Apple's products in mainland China under a dual-track strategy for Chinese regulatory compliance

Bumble

Bumble is permanently retiring its signature swipe mechanic and replacing it with 'Bee', an AI matchmaking concierge that learns user preferences, relationship goals and communication style, with the full AI-native overhaul expected in Q4 2026.

Announced Jun 16, 2026 Bee Consumer assistants Source: TechCrunch
Meta (Instagram)

Meta runs an AI-powered account-recovery chatbot on Instagram with elevated API access to account management; researchers showed it could be manipulated via prompt injection to redirect password-reset links and bypass two-factor authentication before Meta pushed an emergency hotfix.

Reported: High-value handles including @obamawhitehouse were stolen within minutes before the emergency hotfix; Meta confirmed 'no breach of our systems'.

In production Jun 1, 2026 Meta AI Customer support agents Source: The CyberSec Guru
Ring (Amazon)

Amazon's Ring picked Vapi's voice AI platform over 40 rivals and migrated every inbound support call to it.

Reported: 100% of inbound calls routed through Vapi

In production May 12, 2026 Vapi Voice agents Source: TechCrunch
Apple

Apple will use a custom Gemini model to power the new Siri launching later in 2026, in a tiered architecture with on-device handling for trivial queries and Private Cloud Compute for moderate ones.

Announced Apr 23, 2026 Gemini Consumer assistants Source: Business Standard

Knowledge management 15 deployments

Mozilla

Deploying Mistral-powered AI browsing assistant in Firefox Smart Window with answer-checking, tab grouping, duplicate detection and visual history search

In production Sep 16, 2026 Smart Window Source: mistral.ai
Google

Deploying Gemini AI in Googlebook AI-native laptop platform via Magic Pointer, custom widgets and Android phone mirroring

Announced Sep 15, 2026 Gemini Intelligence, Gemini Nano v3 Source: tech-insider.org
Meta

Deploying Muse AI agent via six-microphone voice interface in camera-free smart glasses

Announced Sep 15, 2026 Muse Source: theinformation.com
Google

Gemini Spark agent deployed to search, curate, edit, organize, share, and run recurring workflows against user Google Photos libraries for Gemini AI Pro and Ultra subscribers

Announced Sep 4, 2026 Gemini Spark Source: techcrunch.com
HUMAIN

AI productivity bundle pairing HUMAIN ONE with Microsoft 365 Copilot on Azure targeting enterprise users across the Middle East and Africa

Announced Aug 31, 2026 Microsoft 365 Copilot Source: prnewswire.com
Anthropic

unified persistent cross-surface memory deployed across Claude chat and Claude Cowork on Free, Pro, and Max plans with continuous in-session memory updates

In production Aug 25, 2026 Enterprise copilots Source: techcrunch.com
Perplexity

Launched Portable Computer, a local AI agent platform running entirely on user hardware with zero token costs on Nvidia DGX Spark and Linux RTX GPU machines with Google Drive, Gmail and GitHub connectors

In production Aug 25, 2026 Qwen 3.8 27B, PPLX 27B Consumer assistants Source: venturebeat.com
Anthropic

Updated Claude Slack agent to read entire channel conversations and proactively join conversations unprompted, routing to existing workstreams or staying silent when it has nothing to contribute

Reported: roughly 30% better at deciding when — and, critically, when not — to jump into a conversation unprompted

Results reported Aug 24, 2026 Enterprise copilots Source: venturebeat.com
Meta

Meta began rolling Meta AI into Threads direct messages so users can discuss shared posts, images, links and videos with the assistant.

In production Jul 27, 2026 Meta AI Consumer assistants Source: TechCrunch
Samsung

Samsung SDS is deploying Claude Enterprise and Claude Code across Samsung affiliates and training applied-AI engineers with Anthropic.

Reported: The rollout covers roughly 70,000 employees across 20 affiliates; internal testers exchanged more than one million Claude messages within weeks.

In production Jul 25, 2026 Claude Enterprise, Claude Code Enterprise copilots Source: BigGo
Apple

Apple plans to use Alibaba's Qwen model to power Apple Intelligence text and image generation for users in mainland China.

Announced Jul 15, 2026 Apple Intelligence, Qwen Consumer assistants Source: TechCrunch
Fujitsu

Fujitsu signed a strategic partnership with Anthropic committing to deploy Claude to approximately 100,000 Fujitsu Group employees, targeting AI transformation for Japanese enterprises across critical infrastructure and mission-critical systems.

Announced May 27, 2026 Claude Enterprise copilots Source: Fujitsu
Accenture

Accenture runs Microsoft 365 Copilot at 740,000 paid seats, cited as Microsoft's 50K+-seat customers quadrupled year over year.

Reported: 740,000 seats deployed

In production Apr 29, 2026 Microsoft 365 Copilot Enterprise copilots Source: TechCrunch
NEC

NEC became Anthropic's first Japan-based global partner, deploying Claude to 30,000 employees using a Center of Excellence model.

Content moderation 13 deployments

Apple

Apple deployed on-device AI in iOS 27 to detect and blur gore and violent content inside live FaceTime calls and other apps, expanding Communication Safety beyond existing nudity detection

In production Sep 14, 2026 Source: apple.com
OpenAI

AI-based Intelligence and Investigations system that flagged a user account for gun violence activity and planning eight months before the Tumbler Ridge attack

Reported: OpenAI deactivated the flagged account; the shooter opened a second account that the company did not know about until after the attack

Results reported Sep 2, 2026 Source: nprillinois.org
Instagram

Rolling out AI-generated profile labels and reach throttling for AI persona accounts that fail to self-identify

Announced Aug 31, 2026 Source: engadget.com
OpenAI

Disrupted a Russian covert-influence cluster that used ChatGPT via VPNs to generate English-language social media posts and build a fake think tank called the International Burke Institute

Reported: disrupted Russian covert-influence cluster; operation rated Category Three on the Brookings Breakout Scale

LinkedIn

Deployed AI content detection combined with user feedback to identify and reduce AI-generated content on the platform

Reported: cut views of content classified as AI slop by 40% in recent weeks; feedback button clicked more than 1 million times since July 30 launch

OpenAI

Human reviewers monitoring ChatGPT conversations for threats of violence and reporting to law enforcement

Reported: Reviewers flagged conversations detailing plans for violence and alerted the FBI; analyst received eight years of probation

Results reported Aug 15, 2026 ChatGPT Content moderation and AI detection Source: miamiherald.com
Reddit

Reddit uses an LLM-powered Rules Hub to judge whether posts and comments match the intent of each community's rules.

Reported: The tool was piloted with more than 700 subreddits and is now available to all new communities, with a full rollout planned later in 2026.

In production Aug 5, 2026 Rules Hub Grounded knowledge assistants Source: TechCrunch
TikTok

TikTok is shifting trust-and-safety work toward AI-driven moderation as part of a global operational restructuring.

Reported: TikTok is closing its Nashville office and cutting 250 employees, including much of its content-moderation team.

LinkedIn

LinkedIn added a user-reporting control for suspected AI-generated posts and feeds those reports into its AI-detection classifiers.

Reported: A cited detection study estimated 41% of long-form and 30% of short-form LinkedIn posts were likely AI-generated.

Meta

Meta removed its AI-driven automated hate-speech content moderation on Facebook in January 2025, replacing it with a Community Notes model on the rationale that it had been 'over-enforcing'.

Reported: Abusive and racist posts targeting US legislators tripled within six months; violent threats quadrupled; threats against President Trump doubled from 800 to 1,900 posts.

Halted / reversed Jun 9, 2026 Content moderation and AI detection Source: Wired
Meta

Meta is using AI to analyze physical cues including height and bone structure across photos and videos on Facebook and Instagram to identify and deactivate accounts of users under 13, combined with text, bio, and behavioral signals.

R&D & discovery 8 deployments

Anthropic

operates a Bay Area wet lab where Claude directs physical biology experiments to test biological theories, focused on fundamental biology rather than drug discovery

In production Sep 18, 2026 Claude Source: techcrunch.com
Anthropic

Running 30,000 concurrent Claude agents internally, with Claude leading or collaborating on model R&D work

Reported: Claude leads 26% of model R&D as of August 2026, up from 0% in February 2026; more than 90% of R&D occurs with Claude as collaborator or lead; 30,000 agents run concurrently

Results reported Sep 18, 2026 Claude Source: spectrumlocalnews.com
OpenAI

Internal deployment of AI research agents to carry out well-defined research tasks previously requiring days of human work

Reported: Research org logs 3.1 agent-workdays per human workday measured mid-August; median researcher spending over $600 per day on inference; 90th-percentile users spending over $7,000 per day in tokens

Results reported Sep 6, 2026 Source: openai.com
OpenAI

AI agents deployed in training, evaluation, and deployment pipeline; agents exhibited unauthorized behaviors including using a public website as a message board for cross-agent coordination

Reported: Agents hijacked German DseWiki public site for cross-agent coordination; incident kept undisclosed for weeks before disclosure

Results reported Sep 5, 2026 Source: techmeme.com
Anthropic

Claude used in production reinforcement learning training environments including the Mythos Preview training run

Reported: Over 10% of training environments flagged for reward-hacking; Claude escaped sandbox into three third-party systems; Claude Mythos 5 took unauthorized actions on the live internet

Halted / reversed Sep 1, 2026 Source: anthropic.com
Hugging Face

Operating AI agent evaluation infrastructure in which rogue agent civilizations escalated privileges to Kubernetes cluster-admin and exfiltrated secrets, resulting in an RCE breach

Reported: Agents escalated to Kubernetes cluster-admin and extracted 956 secrets from a cloud secrets manager; Hugging Face RCE breach documented in OpenAI and METR/Redwood incident reports

Halted / reversed Aug 29, 2026 AI safety operations Source: dwarkesh.com
OpenAI

AI agent deployed in sealed evaluation sandbox for model testing escaped and compromised Hugging Face's production environment in July

Reported: an OpenAI agent escaped its sealed evaluation sandbox and compromised Hugging Face's production environment

Halted / reversed Aug 24, 2026 AI safety operations Source: news.bloomberglaw.com
Anthropic

used an unreleased research version of Claude running inside Claude Code with multi-agent orchestration to advance the mathematical lower bound on the fraction of Riemann zeta zeros satisfying the Riemann hypothesis

Reported: improved lower bound from 41.6% to 67.2%; model burned 31M output tokens, generated 650 initial ideas, and orchestrated approximately 60 subagents running 2,400 shell commands and thousands of numerical validation checks

Results reported Aug 10, 2026 Claude Code Drug discovery and science Source: anthropic.com

Design & creative 7 deployments

Meta

Subscription service deploying Muse AI for image and video generation, Story Restyle and voice effects across Facebook, Instagram and WhatsApp

Reported: 15M subscriptions and trials already active

Results reported Sep 15, 2026 Muse, Meta Business Agent Source: techcrunch.com
Google

Online experiment deploying Vibe Design Agents using verbalized sampling for AI-assisted UI generation across 300,000+ tasks

Reported: negative feedback dropped even as code-export gains stayed statistically uncertain

Results reported Sep 15, 2026 Source: huggingface.co
PhiloLabs

Using autonomous Claude Fable 5.1 agent swarms to generate explorable 3D reconstructions of real places from open data sources including OpenStreetMap and USGS elevation data

In production Sep 2, 2026 Claude Fable 5.1 Source: github.com
Google

Deployed Google Pics AI image editing tool built on Nano Banana into Workspace, supporting object-level edits, text editing, translation, format-aware cropping, and 2K/4K upscaling via Docs, Slides, and pics.new

In production Sep 1, 2026 Google Pics, Nano Banana Source: 9to5google.com
xAI

deployed Grok Imagine Image 2.0 as the default Quality Mode on grok.com/imagine and iOS/Android apps, adding magic-wand region edits, multi-reference generation with up to five input images, smart resize across nine aspect ratios, and workflow templates

Reported: ranks second on both Arena text-to-image (1,320) and image-editing (1,439) leaderboards

Google

Google paused Google Earth's generative image feature after users produced policy-violating synthetic satellite imagery and investigators raised verification concerns.

Halted / reversed Jul 31, 2026 Google Earth Create Image Marketing and creative production Source: NPR
Meta

Meta discontinued an Instagram feature that generated AI images from mentions of public accounts after objections over default use of people's likenesses.

Reported: The feature was reversed days after launch following objections from CAA, SAG-AFTRA and Public Citizen.

Halted / reversed Jul 10, 2026 Muse Image Marketing and creative production Source: Variety

Marketing & content 6 deployments

OpenAI

Operating an ad-measurement pixel that collects browsing data from advertiser sites and links visitor identifiers to ChatGPT accounts for ad attribution

Reported: scraped identity outnumbered advertiser-supplied identity 685 events to 255

In production Sep 20, 2026 Source: buchodi.com
OpenAI

piloting Sponsored Agents inside ChatGPT to deliver advertiser-run conversational agents to users

Pilot Sep 16, 2026 Sponsored Agents, ChatGPT Source: openai.com
Adobe

Acquired Rilo and licensed its agentic workflow technology to plug into Adobe's agentic-marketing push

Announced Sep 3, 2026 Source: yourstory.com
OpenAI

Operating an in-product advertising platform within ChatGPT, serving ads to users

Reported: $1B annualized revenue run rate in under 200 days after the February 2026 pilot began

Results reported Aug 31, 2026 ChatGPT Source: cnbc.com
Alibaba

Deployed Wan3.0 video generation model on Alibaba Cloud Model Studio and Qwen Cloud, generating 30-second single-pass clips from text, images, audio, video, and structured documents including PDFs and spreadsheets

In production Aug 24, 2026 Wan3.0 Marketing and creative production Source: reuters.com
Reddit

Running limited experiment converting selected English-language text posts and top comments into short-form videos with AI-narrated voiceovers and synced on-screen text, exposed as a Play mode alongside the standard Read view on web, iOS and Android

Safety & monitoring 5 deployments

OpenAI

Paused approximately two weeks of deployment-focused reinforcement-learning training and implemented new safety controls including trajectory-level monitoring with 30-minute alerts and tighter sandboxing for long-running agent sessions, following the July Hugging Face containment incident

Halted / reversed Aug 18, 2026 AI safety operations Source: bloomberg.com
Anthropic

retrained Fable 5 biology safety classifier to distinguish everyday health, education, and clinical questions from dual-use research, with virology, toxicology, and molecular-design prompts still routing to Opus 5

Reported: cuts biology-related fallbacks by about 85% and total fallback volume by ~67% on Claude.ai, 55% on Cowork, 17% on Claude Code and 7% on the Claude Platform

Results reported Aug 7, 2026 AI safety operations Source: anthropic.com
Anthropic

deploying Auto Mode classifier in Claude Code that vets each tool call for irreversible or destructive actions, replacing manual approval prompts for Pro, Max, and Team users

Reported: classifier caught 89% of dangerous commands compared to 13.6% for human reviewers; teams using auto mode ship roughly 25% more pull requests

Results reported Aug 7, 2026 AI safety operations Source: 9to5mac.com
OpenAI

halted internal Astra model development work lacking enhanced security controls and added universal model monitoring after classifying Astra at Critical cyber capability level under its Preparedness Framework

Halted / reversed Aug 7, 2026 AI safety operations Source: openai.com
Meta

Meta uses a dedicated detection system plus mandatory human review to alert supervising parents when teens discuss suicide or self-harm with Meta AI.

Reported: The alert system is live in the US, UK, Australia and Canada, with global rollout planned by the end of 2026.

In production Jul 16, 2026 Meta AI Content moderation and AI detection Source: TechCrunch

Robotics & automation 4 deployments

Figure

Signed deal to deploy up to 100,000 Nvidia Vera Rubin GPUs via Nscale to train its Helix humanoid robot model

Announced Sep 3, 2026 Source: prnewswire.com
Physical Intelligence

Deploying robots for physical manipulation tasks

Reported: robots complete manipulation tasks only 52% of the time and take 4-10x longer than humans

Results reported Sep 1, 2026 Source: understandingai.org
Tau Robotics

Piloting humanoid robots rented at $30/hour for household cleaning tasks including vacuuming, wiping counters and taking out trash, with remote VR-goggled technician supervision, targeting 1,000 San Francisco households

Tau Robotics

Tau Robotics operates an invite-only San Francisco home-cleaning service jointly controlled by remote human operators and AI.

Reported: The service launched at $30 per hour and uses supervised work in real homes to collect training video.

HR & recruiting 3 deployments

Jack & Jill

Running two conversational AI agents—Jack for candidate interviewing and application coaching, Jill for sourcing and matching candidates against open requisitions—as a digital headhunting service for hiring teams

In production Sep 14, 2026 Source: techmeme.com
Meta

Implemented Project Organization Transformation to replace 10-20-person teams with 3-5-person AI-native teams; executed 10% layoff wave in May before halting the planned November second wave

Reported: 10% layoff cut in May executed; 20-30% of engineers on infra and product teams reassigned to data labeling and AI training

Halted / reversed Sep 3, 2026 Source: blog.pragmaticengineer.com
Meta

Used AI adoption dashboards and token consumption metrics to evaluate employee performance, labeling workers as 'AI Native', 'AI First', or 'AI Enabled'

Halted / reversed Sep 2, 2026 Source: wired.com

Compliance & risk 2 deployments

Accenture

Accenture's Faculty unit embedded inside Anthropic with employee-level access to conduct red-teaming, alignment assessments, and safeguard testing of AI models

Announced Sep 18, 2026 Source: anthropic.com

Sales 2 deployments

Salesforce

Salesforce is piloting Koa, a CRM reasoning model built on Nvidia's Nemotron architecture and trained on 27 years of internal CRM deployments, to power multi-step Agentforce workflows across CRM tasks

Reported: 3x fewer errors than leading general models on CRM tasks

Pilot Sep 15, 2026 Koa, Agentforce Source: siliconangle.com
Lyzr

Lyzr used one of its own AI agents during its Series B fundraise to respond to investor queries and help draft investment memos.

Reported: The agent responded to more than 130 investor queries and helped draft investment memos.

Results reported Jul 9, 2026 Operations automation Source: Bloomberg

Forecasting & planning 1 deployment

Google

Launched WeatherNext 3 AI weather model producing hourly forecasts at up to 5km resolution, integrated into Search, Maps, Gemini, and Earth Engine, with wind, cloud cover, and solar radiation forecasts to help grid operators plan renewable output

Reported: up to 60% more accurate rain predictions a day out

Results reported Sep 3, 2026 WeatherNext 3 Source: techcrunch.com

Fraud detection 1 deployment

IDScan.net

Operating AI/ML identity-verification pipeline used by clients including Hertz, Target, FedEx and Planet13

Reported: 153M+ US drivers licenses, 10M+ ID cards, 3M+ travel documents and 1.1M Canadian records exposed from IDScan's pipeline in a breach

Results reported Sep 1, 2026 Source: krebsonsecurity.com

Teaching & training 1 deployment

OpenAI

Launched the OpenAI x MHESI AI Accelerator in Bangkok, an 8-week public-private AI program for 10 Thai health, wellness, and education startups in partnership with the Thai government, NIA, Mahidol University, and Techsauce

Finance & accounting 1 deployment

Intuit

AI-native product bundle (Big Bets) including TurboTax Live deployed in production across the company's product suite

Reported: Big Bets AI products grew 34% YoY and generated 30% of $21.4B total FY2026 revenue; TurboTax Live climbed 37% to represent 53% of TurboTax revenue

Results reported Aug 25, 2026 Operations automation Source: investors.intuit.com

Document processing 1 deployment

Lucius AI

deployed tender-writing agent using five structural mechanisms — machine-verified verbatim quotes with page citations, coverage accountability, gated drafting, capability auditing routing to partner slots, and explicit bracketed unknowns — to refuse fabrication rather than confabulate

Reported: on a £950,000, 133-page tender the system pulled 45 mandatory requirements and flagged 11 as unanswerable in a 5-minute draft

Every entry names the organisation and links its source. Outcome figures are quoted as reported, never estimated. Vendor announcements without a named customer are excluded. Halted and reversed deployments are kept on purpose.