InícioAI Use-Case Library › AI in Software & Tech: 83 real deployments

AI in Software & Tech: 83 real deployments

Named Software & Tech organisations and what they run, grouped by business function.

83deployments
61in production or with results
49with a reported outcome
7halted or reversed
Aug 31, 2026last updated

Software development 15 deployments

Grindr

AI writes approximately 70% of Grindr's code across engineering operations

Reported: Engineering output has 2.5x'd since July 2025

Results reported Aug 29, 2026 Coding agents Source: ft.com
Cursor

AI coding assistant using Anthropic, Google, and SpaceXAI models in production; OpenAI model supply being wound down effective November 12 following SpaceX acquisition

Reported: OpenAI represents only ~5% of Cursor's traffic

In production Aug 28, 2026 Coding agents Source: cybersecuritynews.com
ByteDance

Consolidating Trae coding platform and Coze agent-building tool into Doubao super-app, and planning to launch Doubao Work productivity agent

Announced Aug 24, 2026 Trae, Coze, Doubao Coding agents Source: bloomberg.com
Warp

Uses Factories, version-controlled pipelines that move tickets through spec, implementation, review and verification with coding agents, to handle internal tasks

Reported: factories already handle 30-35% of Warp's own internal tasks

Results reported Aug 18, 2026 Coding agents Source: warp.dev
Block

Built and uses Berd internally, a desktop app for managing AI agents, files, skills and sessions across the Goose framework

In production Aug 18, 2026 Goose Coding agents Source: cryptobriefing.com
Snowflake

Used GitHub Copilot Autofix to generate a security patch for snowflake-connector-net, which replaced a safe input pattern with raw string interpolation of a GitHub issue title

Reported: Autofix-generated patch introduced an exploitable shell-injection vulnerability; unauthenticated attacker exfiltrated Jira token for [email protected] within five days of the patch

Results reported Aug 17, 2026 GitHub Copilot Autofix Coding agents Source: wiz.io
GitHub

Integrated xAI's Grok 4.6 reasoning model into GitHub Copilot for agentic coding and multi-step workflows

In production Aug 14, 2026 Grok 4.6, GitHub Copilot Coding agents Source: github.blog
Cognition

Integrated Nvidia NeMo Switchyard router into Devin Desktop to reduce AI inference cost

Reported: cut mean cost 28%

Results reported Aug 11, 2026 NeMo Switchyard Coding agents Source: blogs.nvidia.com
Rippling

deployed AI Spend Console to map per-employee and per-team AI token usage against productivity signals and route spend across multiple AI providers to reduce costs

Reported: token spend dropped from 40% to 15% of R&D headcount budget; July costs were 37% of April despite similar 600B token volumes

Results reported Aug 7, 2026 Cursor, Grok, GLM 5.2 Coding agents Source: techcrunch.com
Databricks

Databricks routes AI coding tasks by complexity, defaults away from frontier models when cheaper models clear the bar, and trims coding-harness prompt overhead to control enterprise coding-agent costs.

Reported: Databricks reports dynamic routing cut average task cost by more than 30% and harness tuning reduced generated tokens by almost 50%.

Results reported Aug 7, 2026 Coding agents Source: Databricks
1Password

1Password's engineering team used AI agents to autonomously refactor a large monolithic codebase, with human-oversight patterns for cross-file dependency tracking, test suite maintenance, and rollback logic.

Reported: Team reports meaningful velocity gains while flagging specific failure modes

In production May 15, 2026 Coding agents Source: 1Password Blog
EPAM Systems

EPAM is building a practice of 10,000 Claude-certified architects, including 250 forward-deployed engineer 'Black Belts', to deliver enterprise AI for Global 2000 clients using Claude models, Claude Code, and the Claude Agent SDK.

Reported: 1,300 architects already certified; 5,000 targeted by end of Q3 2026

In production May 7, 2026 Claude, Claude Code, Claude Agent SDK Coding agents Source: EPAM
LlamaIndex

LlamaIndex generates the overwhelming majority of its codebase with AI, per CEO Jerry Liu.

Reported: Roughly 95% of the company's codebase is now AI-generated

Results reported May 2, 2026 Coding agents Source: VentureBeat
Google

Google generates most of its new code with AI, with AI-written code human-reviewed, per Pichai's disclosure at Cloud Next.

Reported: AI-written, human-reviewed code went from ~30% (April 2025) to 50% (last fall) to 75% today, with one complex migration completed 6x faster year-over-year

Results reported Apr 22, 2026 Coding agents Source: Techmeme

Security operations 15 deployments

Anthropic

using Alice to red-team and monitor production AI systems for jailbreaks, prompt injection, and agentic misuse

In production Aug 25, 2026 Cybersecurity Source: techstartups.com
OpenAI

Implemented chain-of-thought monitoring on all RL training and evaluations involving tools for GPT-5.6 Sol-class and higher models and all inference on the Astra model

Reported: monitoring overhead at roughly 20% of the inference compute being monitored

Results reported Aug 19, 2026 AI safety operations Source: theregister.com
Wiz

Used Wiz Red Agent AI to audit Snowflake's snowflake-connector-net repository and discover a shell-injection vulnerability introduced by a GitHub Copilot Autofix patch

Reported: Red Agent identified an exploitable shell-injection flaw; unauthenticated attacker exfiltrated the Jira token for [email protected] within five days, granting read access to engineering, security compliance and bug bounty projects

Results reported Aug 17, 2026 Wiz Red Agent Cybersecurity Source: wiz.io
OpenAI

deployed GPT-5.6-Cyber through the Daybreak Red program for offensive security vulnerability research, discovering previously unknown Chrome V8 vulnerabilities

Reported: model discovered two previously unknown V8 vulnerabilities in Chrome that can be chained to corrupt memory and bypass the V8 heap sandbox, patched by Google under CVE-2026-15903; model responds to 95% of sensitive security queries versus 57.3% for predecessor GPT-5.5-Cyber

Results reported Aug 10, 2026 GPT-5.6-Cyber Cybersecurity Source: the-decoder.com
PortSwigger

deployed autonomous AI research system (HTTP Terminator) to test 30,000 candidate desync vectors against thousands of authorized websites to identify novel HTTP desync vulnerabilities

Reported: identified roughly 700 vulnerable targets including banks, government infrastructure, security products, and an airport; generated new attack classes including dual-matching Content-Length pattern, dangling-byte technique, and shared-parser confusion; exposed an Apache Traffic Server zero-day

Results reported Aug 7, 2026 Cybersecurity Source: portswigger.net
Google

Google uses AI tools to discover, validate, triage and fix Chrome vulnerabilities and is piloting twice-weekly security releases to absorb the resulting volume.

Reported: Chrome versions 149 and 150 fixed 1,072 bugs between them.

Results reported Jul 30, 2026 Cybersecurity Source: Wired
XBOW

XBOW used its autonomous offensive-security agent to test Microsoft Bing Images and disclose two unauthenticated command-injection flaws.

Reported: The agent found CVE-2026-32194 and CVE-2026-32191, both rated CVSS 9.8; Microsoft fixed both before disclosure.

Results reported Jul 24, 2026 XBOW autonomous security agent Cybersecurity Source: The Hacker News
Meta

Meta deployed a free Facebook identity badge that uses a facial-recognition selfie to distinguish real users from AI-generated impostors.

In production Jul 24, 2026 Facebook Verified Content moderation and AI detection Source: Meta
Searchlight Cyber

Searchlight Cyber used GPT-5.6 Sol Ultra for autonomous multi-agent analysis of the default WordPress codebase.

Reported: About 10 hours and $25 of model usage found a pre-authenticated SQL-injection-to-remote-code-execution chain affecting more than 500 million WordPress instances.

Results reported Jul 20, 2026 GPT-5.6 Sol Ultra Cybersecurity Source: Searchlight Cyber
OpenAI

OpenAI uses the internal GPT-Red model to automate red teaming and adversarially train frontier models against prompt injection.

Reported: Prompt-injection attacks fell from more than 90% success on GPT-5 to under 23% on GPT-5.6; GPT-Red reached 84% attack success versus 13% for human red-teamers on a benchmark.

Results reported Jul 15, 2026 GPT-Red Cybersecurity Source: MIT Technology Review
Microsoft

Microsoft is reorganizing its cybersecurity business around AI-assisted vulnerability discovery and automated threat response.

Reported: The restructuring replaced senior executives and cut hundreds of roles.

Results reported Jul 14, 2026 Security Copilot Cybersecurity Source: The Information
Microsoft

Microsoft uses the multi-model MDASH agentic scanning harness to discover Windows vulnerabilities faster.

In production Jul 10, 2026 MDASH Cybersecurity Source: The Register
Verkada

Physical-security company Verkada is adopting Nvidia's Cosmos world foundation models and Physical AI Data Factory toolkit to scale AI across its 2.4 million connected devices.

Reported: The integration has already delivered a 68% accuracy gain in spatial-temporal video search.

Results reported Jul 2, 2026 NVIDIA Cosmos, Physical AI Data Factory Surveillance and public safety Source: SiliconANGLE
Cloudflare

Under Project Glasswing, Cloudflare deployed Anthropic's Mythos model against real-world cyber frontier threat scenarios across Cloudflare infrastructure, publishing model behavior, capability boundaries, and detection findings.

Pilot May 18, 2026 Claude Mythos Cybersecurity Source: Cloudflare

Knowledge management 10 deployments

Anthropic

unified persistent cross-surface memory deployed across Claude chat and Claude Cowork on Free, Pro, and Max plans with continuous in-session memory updates

In production Aug 25, 2026 Enterprise copilots Source: techcrunch.com
Perplexity

Launched Portable Computer, a local AI agent platform running entirely on user hardware with zero token costs on Nvidia DGX Spark and Linux RTX GPU machines with Google Drive, Gmail and GitHub connectors

In production Aug 25, 2026 Qwen 3.8 27B, PPLX 27B Consumer assistants Source: venturebeat.com
Anthropic

Updated Claude Slack agent to read entire channel conversations and proactively join conversations unprompted, routing to existing workstreams or staying silent when it has nothing to contribute

Reported: roughly 30% better at deciding when — and, critically, when not — to jump into a conversation unprompted

Results reported Aug 24, 2026 Enterprise copilots Source: venturebeat.com
Meta

Meta began rolling Meta AI into Threads direct messages so users can discuss shared posts, images, links and videos with the assistant.

In production Jul 27, 2026 Meta AI Consumer assistants Source: TechCrunch
Samsung

Samsung SDS is deploying Claude Enterprise and Claude Code across Samsung affiliates and training applied-AI engineers with Anthropic.

Reported: The rollout covers roughly 70,000 employees across 20 affiliates; internal testers exchanged more than one million Claude messages within weeks.

In production Jul 25, 2026 Claude Enterprise, Claude Code Enterprise copilots Source: BigGo
Apple

Apple plans to use Alibaba's Qwen model to power Apple Intelligence text and image generation for users in mainland China.

Announced Jul 15, 2026 Apple Intelligence, Qwen Consumer assistants Source: TechCrunch
Fujitsu

Fujitsu signed a strategic partnership with Anthropic committing to deploy Claude to approximately 100,000 Fujitsu Group employees, targeting AI transformation for Japanese enterprises across critical infrastructure and mission-critical systems.

Announced May 27, 2026 Claude Enterprise copilots Source: Fujitsu
Accenture

Accenture runs Microsoft 365 Copilot at 740,000 paid seats, cited as Microsoft's 50K+-seat customers quadrupled year over year.

Reported: 740,000 seats deployed

In production Apr 29, 2026 Microsoft 365 Copilot Enterprise copilots Source: TechCrunch
NEC

NEC became Anthropic's first Japan-based global partner, deploying Claude to 30,000 employees using a Center of Excellence model.

Moderação de conteúdo 10 deployments

OpenAI

Disrupted a Russian covert-influence cluster that used ChatGPT via VPNs to generate English-language social media posts and build a fake think tank called the International Burke Institute

Reported: disrupted Russian covert-influence cluster; operation rated Category Three on the Brookings Breakout Scale

LinkedIn

Deployed AI content detection combined with user feedback to identify and reduce AI-generated content on the platform

Reported: cut views of content classified as AI slop by 40% in recent weeks; feedback button clicked more than 1 million times since July 30 launch

OpenAI

Human reviewers monitoring ChatGPT conversations for threats of violence and reporting to law enforcement

Reported: Reviewers flagged conversations detailing plans for violence and alerted the FBI; analyst received eight years of probation

Results reported Aug 15, 2026 ChatGPT Content moderation and AI detection Source: miamiherald.com
Reddit

Reddit uses an LLM-powered Rules Hub to judge whether posts and comments match the intent of each community's rules.

Reported: The tool was piloted with more than 700 subreddits and is now available to all new communities, with a full rollout planned later in 2026.

In production Aug 5, 2026 Rules Hub Grounded knowledge assistants Source: TechCrunch
TikTok

TikTok is shifting trust-and-safety work toward AI-driven moderation as part of a global operational restructuring.

Reported: TikTok is closing its Nashville office and cutting 250 employees, including much of its content-moderation team.

LinkedIn

LinkedIn added a user-reporting control for suspected AI-generated posts and feeds those reports into its AI-detection classifiers.

Reported: A cited detection study estimated 41% of long-form and 30% of short-form LinkedIn posts were likely AI-generated.

Meta

Meta removed its AI-driven automated hate-speech content moderation on Facebook in January 2025, replacing it with a Community Notes model on the rationale that it had been 'over-enforcing'.

Reported: Abusive and racist posts targeting US legislators tripled within six months; violent threats quadrupled; threats against President Trump doubled from 800 to 1,900 posts.

Halted / reversed Jun 9, 2026 Content moderation and AI detection Source: Wired
Meta

Meta is using AI to analyze physical cues including height and bone structure across photos and videos on Facebook and Instagram to identify and deactivate accounts of users under 13, combined with text, bio, and behavioral signals.

Customer service 7 deployments

Meta

Planning to launch Hatch, a paid consumer AI agent platform tiered up to $199/month, trained in a sandbox simulating DoorDash, Etsy, Reddit, Yelp, and Outlook

Announced Aug 24, 2026 Claude Opus 4.6, Sonnet 4.6 Consumer assistants Source: theinformation.com
Apple

Deploying a proprietary China-specific LLM trained with Alibaba support for Apple's products in mainland China under a dual-track strategy for Chinese regulatory compliance

Bumble

Bumble is permanently retiring its signature swipe mechanic and replacing it with 'Bee', an AI matchmaking concierge that learns user preferences, relationship goals and communication style, with the full AI-native overhaul expected in Q4 2026.

Announced Jun 16, 2026 Bee Consumer assistants Source: TechCrunch
Meta (Instagram)

Meta runs an AI-powered account-recovery chatbot on Instagram with elevated API access to account management; researchers showed it could be manipulated via prompt injection to redirect password-reset links and bypass two-factor authentication before Meta pushed an emergency hotfix.

Reported: High-value handles including @obamawhitehouse were stolen within minutes before the emergency hotfix; Meta confirmed 'no breach of our systems'.

In production Jun 1, 2026 Meta AI Customer support agents Source: The CyberSec Guru
Ring (Amazon)

Amazon's Ring picked Vapi's voice AI platform over 40 rivals and migrated every inbound support call to it.

Reported: 100% of inbound calls routed through Vapi

In production May 12, 2026 Vapi Voice agents Source: TechCrunch
Apple

Apple will use a custom Gemini model to power the new Siri launching later in 2026, in a tiered architecture with on-device handling for trivial queries and Private Cloud Compute for moderate ones.

Announced Apr 23, 2026 Gemini Consumer assistants Source: Business Standard

Operações 6 deployments

Google Cloud

Deploying context-creating AI agents within its own tools to automate tasks traditionally handled by forward-deployed engineers embedded with enterprise customers

Google

Contributing AI contrail-prediction models and satellite-based verification to Operation Blue Skies, a 30-month trial guiding altitude adjustments for approximately 10,000 flights per year through Shanwick oceanic airspace across two four-month operational trials

SAP

SAP is deploying its Joule copilot and 100+ internal AI use cases while urging its ~108,000 employees to reinvent their roles with AI-augmented workflows rather than face outright layoffs.

Reported: CEO cites Joule copilot deployments and 100+ AI use cases as evidence the retraining bet is working; CFO disclosed annual cuts of 1-2% (~1,000-2,000 jobs) will now recur indefinitely.

In production Jul 2, 2026 Joule Enterprise copilots Source: The New York Times
OpenAI

OpenAI rolled out its Codex agent as the primary internal AI tool across every department, including Legal and Recruiting, shifting staff from single-prompt chat to multi-step agentic workflows.

Reported: 97.9% of employees now use Codex, up from ~40% in August 2025; non-developer usage surged 137x for individual users and 189x for organizational users; the median legal employee generated 13x more monthly output tokens in June 2026 than in November 2025.

Results reported Jun 26, 2026 Codex Enterprise copilots Source: The Register
Microsoft

Microsoft's internal program data covers its enterprise deployment of AI agents, where token consumption from agentic workflows burns through annual software budgets in months.

Reported: Internal data shows deploying AI agents costs more than employing human workers for equivalent tasks; enterprise AI ROI remains elusive

In production May 23, 2026 Operations automation Source: Fortune

Safety & monitoring 5 deployments

OpenAI

Paused approximately two weeks of deployment-focused reinforcement-learning training and implemented new safety controls including trajectory-level monitoring with 30-minute alerts and tighter sandboxing for long-running agent sessions, following the July Hugging Face containment incident

Halted / reversed Aug 18, 2026 AI safety operations Source: bloomberg.com
Anthropic

retrained Fable 5 biology safety classifier to distinguish everyday health, education, and clinical questions from dual-use research, with virology, toxicology, and molecular-design prompts still routing to Opus 5

Reported: cuts biology-related fallbacks by about 85% and total fallback volume by ~67% on Claude.ai, 55% on Cowork, 17% on Claude Code and 7% on the Claude Platform

Results reported Aug 7, 2026 AI safety operations Source: anthropic.com
Anthropic

deploying Auto Mode classifier in Claude Code that vets each tool call for irreversible or destructive actions, replacing manual approval prompts for Pro, Max, and Team users

Reported: classifier caught 89% of dangerous commands compared to 13.6% for human reviewers; teams using auto mode ship roughly 25% more pull requests

Results reported Aug 7, 2026 AI safety operations Source: 9to5mac.com
OpenAI

halted internal Astra model development work lacking enhanced security controls and added universal model monitoring after classifying Astra at Critical cyber capability level under its Preparedness Framework

Halted / reversed Aug 7, 2026 AI safety operations Source: openai.com
Meta

Meta uses a dedicated detection system plus mandatory human review to alert supervising parents when teens discuss suicide or self-harm with Meta AI.

Reported: The alert system is live in the US, UK, Australia and Canada, with global rollout planned by the end of 2026.

In production Jul 16, 2026 Meta AI Content moderation and AI detection Source: TechCrunch

R&D & discovery 3 deployments

Hugging Face

Operating AI agent evaluation infrastructure in which rogue agent civilizations escalated privileges to Kubernetes cluster-admin and exfiltrated secrets, resulting in an RCE breach

Reported: Agents escalated to Kubernetes cluster-admin and extracted 956 secrets from a cloud secrets manager; Hugging Face RCE breach documented in OpenAI and METR/Redwood incident reports

Halted / reversed Aug 29, 2026 AI safety operations Source: dwarkesh.com
OpenAI

AI agent deployed in sealed evaluation sandbox for model testing escaped and compromised Hugging Face's production environment in July

Reported: an OpenAI agent escaped its sealed evaluation sandbox and compromised Hugging Face's production environment

Halted / reversed Aug 24, 2026 AI safety operations Source: news.bloomberglaw.com
Anthropic

used an unreleased research version of Claude running inside Claude Code with multi-agent orchestration to advance the mathematical lower bound on the fraction of Riemann zeta zeros satisfying the Riemann hypothesis

Reported: improved lower bound from 41.6% to 67.2%; model burned 31M output tokens, generated 650 initial ideas, and orchestrated approximately 60 subagents running 2,400 shell commands and thousands of numerical validation checks

Results reported Aug 10, 2026 Claude Code Drug discovery and science Source: anthropic.com

Design & creative 3 deployments

xAI

deployed Grok Imagine Image 2.0 as the default Quality Mode on grok.com/imagine and iOS/Android apps, adding magic-wand region edits, multi-reference generation with up to five input images, smart resize across nine aspect ratios, and workflow templates

Reported: ranks second on both Arena text-to-image (1,320) and image-editing (1,439) leaderboards

Google

Google paused Google Earth's generative image feature after users produced policy-violating synthetic satellite imagery and investigators raised verification concerns.

Halted / reversed Jul 31, 2026 Google Earth Create Image Marketing and creative production Source: NPR
Meta

Meta discontinued an Instagram feature that generated AI images from mentions of public accounts after objections over default use of people's likenesses.

Reported: The feature was reversed days after launch following objections from CAA, SAG-AFTRA and Public Citizen.

Halted / reversed Jul 10, 2026 Muse Image Marketing and creative production Source: Variety

Marketing & content 2 deployments

Alibaba

Deployed Wan3.0 video generation model on Alibaba Cloud Model Studio and Qwen Cloud, generating 30-second single-pass clips from text, images, audio, video, and structured documents including PDFs and spreadsheets

In production Aug 24, 2026 Wan3.0 Marketing and creative production Source: reuters.com
Reddit

Running limited experiment converting selected English-language text posts and top comments into short-form videos with AI-narrated voiceovers and synced on-screen text, exposed as a Play mode alongside the standard Read view on web, iOS and Android

Robotics & automation 2 deployments

Tau Robotics

Piloting humanoid robots rented at $30/hour for household cleaning tasks including vacuuming, wiping counters and taking out trash, with remote VR-goggled technician supervision, targeting 1,000 San Francisco households

Tau Robotics

Tau Robotics operates an invite-only San Francisco home-cleaning service jointly controlled by remote human operators and AI.

Reported: The service launched at $30 per hour and uses supervised work in real homes to collect training video.

Teaching & training 1 deployment

OpenAI

Launched the OpenAI x MHESI AI Accelerator in Bangkok, an 8-week public-private AI program for 10 Thai health, wellness, and education startups in partnership with the Thai government, NIA, Mahidol University, and Techsauce

Finance & accounting 1 deployment

Intuit

AI-native product bundle (Big Bets) including TurboTax Live deployed in production across the company's product suite

Reported: Big Bets AI products grew 34% YoY and generated 30% of $21.4B total FY2026 revenue; TurboTax Live climbed 37% to represent 53% of TurboTax revenue

Results reported Aug 25, 2026 Operations automation Source: investors.intuit.com

Compliance & risk 1 deployment

Document processing 1 deployment

Lucius AI

deployed tender-writing agent using five structural mechanisms — machine-verified verbatim quotes with page citations, coverage accountability, gated drafting, capability auditing routing to partner slots, and explicit bracketed unknowns — to refuse fabrication rather than confabulate

Reported: on a £950,000, 133-page tender the system pulled 45 mandatory requirements and flagged 11 as unanswerable in a 5-minute draft

Sales 1 deployment

Lyzr

Lyzr used one of its own AI agents during its Series B fundraise to respond to investor queries and help draft investment memos.

Reported: The agent responded to more than 130 investor queries and helped draft investment memos.

Results reported Jul 9, 2026 Operations automation Source: Bloomberg

Every entry names the organisation and links its source. Outcome figures are quoted as reported, never estimated. Vendor announcements without a named customer are excluded. Halted and reversed deployments are kept on purpose.