HomeAI Use-Case Library › AI in security operations: 22 real deployments

AI in security operations: 22 real deployments

Named deployments in security operations, grouped by industry.

22deployments
17in production or with results
11with a reported outcome
2halted or reversed
Aug 30, 2026last updated

Software & Tech 15 deployments

Anthropic

using Alice to red-team and monitor production AI systems for jailbreaks, prompt injection, and agentic misuse

In production Aug 25, 2026 Cybersecurity Source: techstartups.com
OpenAI

Implemented chain-of-thought monitoring on all RL training and evaluations involving tools for GPT-5.6 Sol-class and higher models and all inference on the Astra model

Reported: monitoring overhead at roughly 20% of the inference compute being monitored

Results reported Aug 19, 2026 AI safety operations Source: theregister.com
Wiz

Used Wiz Red Agent AI to audit Snowflake's snowflake-connector-net repository and discover a shell-injection vulnerability introduced by a GitHub Copilot Autofix patch

Reported: Red Agent identified an exploitable shell-injection flaw; unauthenticated attacker exfiltrated the Jira token for [email protected] within five days, granting read access to engineering, security compliance and bug bounty projects

Results reported Aug 17, 2026 Wiz Red Agent Cybersecurity Source: wiz.io
OpenAI

deployed GPT-5.6-Cyber through the Daybreak Red program for offensive security vulnerability research, discovering previously unknown Chrome V8 vulnerabilities

Reported: model discovered two previously unknown V8 vulnerabilities in Chrome that can be chained to corrupt memory and bypass the V8 heap sandbox, patched by Google under CVE-2026-15903; model responds to 95% of sensitive security queries versus 57.3% for predecessor GPT-5.5-Cyber

Results reported Aug 10, 2026 GPT-5.6-Cyber Cybersecurity Source: the-decoder.com
PortSwigger

deployed autonomous AI research system (HTTP Terminator) to test 30,000 candidate desync vectors against thousands of authorized websites to identify novel HTTP desync vulnerabilities

Reported: identified roughly 700 vulnerable targets including banks, government infrastructure, security products, and an airport; generated new attack classes including dual-matching Content-Length pattern, dangling-byte technique, and shared-parser confusion; exposed an Apache Traffic Server zero-day

Results reported Aug 7, 2026 Cybersecurity Source: portswigger.net
Google

Google uses AI tools to discover, validate, triage and fix Chrome vulnerabilities and is piloting twice-weekly security releases to absorb the resulting volume.

Reported: Chrome versions 149 and 150 fixed 1,072 bugs between them.

Results reported Jul 30, 2026 Cybersecurity Source: Wired
XBOW

XBOW used its autonomous offensive-security agent to test Microsoft Bing Images and disclose two unauthenticated command-injection flaws.

Reported: The agent found CVE-2026-32194 and CVE-2026-32191, both rated CVSS 9.8; Microsoft fixed both before disclosure.

Results reported Jul 24, 2026 XBOW autonomous security agent Cybersecurity Source: The Hacker News
Meta

Meta deployed a free Facebook identity badge that uses a facial-recognition selfie to distinguish real users from AI-generated impostors.

In production Jul 24, 2026 Facebook Verified Content moderation and AI detection Source: Meta
Searchlight Cyber

Searchlight Cyber used GPT-5.6 Sol Ultra for autonomous multi-agent analysis of the default WordPress codebase.

Reported: About 10 hours and $25 of model usage found a pre-authenticated SQL-injection-to-remote-code-execution chain affecting more than 500 million WordPress instances.

Results reported Jul 20, 2026 GPT-5.6 Sol Ultra Cybersecurity Source: Searchlight Cyber
OpenAI

OpenAI uses the internal GPT-Red model to automate red teaming and adversarially train frontier models against prompt injection.

Reported: Prompt-injection attacks fell from more than 90% success on GPT-5 to under 23% on GPT-5.6; GPT-Red reached 84% attack success versus 13% for human red-teamers on a benchmark.

Results reported Jul 15, 2026 GPT-Red Cybersecurity Source: MIT Technology Review
Microsoft

Microsoft is reorganizing its cybersecurity business around AI-assisted vulnerability discovery and automated threat response.

Reported: The restructuring replaced senior executives and cut hundreds of roles.

Results reported Jul 14, 2026 Security Copilot Cybersecurity Source: The Information
Microsoft

Microsoft uses the multi-model MDASH agentic scanning harness to discover Windows vulnerabilities faster.

In production Jul 10, 2026 MDASH Cybersecurity Source: The Register
Verkada

Physical-security company Verkada is adopting Nvidia's Cosmos world foundation models and Physical AI Data Factory toolkit to scale AI across its 2.4 million connected devices.

Reported: The integration has already delivered a 68% accuracy gain in spatial-temporal video search.

Results reported Jul 2, 2026 NVIDIA Cosmos, Physical AI Data Factory Surveillance and public safety Source: SiliconANGLE
Cloudflare

Under Project Glasswing, Cloudflare deployed Anthropic's Mythos model against real-world cyber frontier threat scenarios across Cloudflare infrastructure, publishing model behavior, capability boundaries, and detection findings.

Pilot May 18, 2026 Claude Mythos Cybersecurity Source: Cloudflare

Government & Public Sector 3 deployments

Motor Vehicle Crime Prevention Authority

Installed more than 3,000 AI-powered license plate reader cameras via grants to cities and counties since 2023 for vehicle surveillance

Reported: Texas Gov. Abbott ordered state agencies to pause funding amid privacy concerns and misuse reports; cities and counties canceling contracts; Lufkin officer facing 100 counts for surveilling 11 people

Halted / reversed Aug 29, 2026 Surveillance and public safety Source: foxnews.com
ICE

Agents wore Meta's AI smart glasses during immigration enforcement operations across six states; internal memo now bars employees from wearing the devices due to risk of capturing sensitive data

Halted / reversed Aug 18, 2026 Surveillance and public safety Source: nytimes.com
US Treasury, CISA, DHS and Department of Defense

The Gold Eagle federal clearinghouse uses frontier AI to scan government and private-sector systems for vulnerabilities and prioritize patches.

In production Jul 14, 2026 Gold Eagle Cybersecurity Source: CyberScoop

Military 2 deployments

U.S. Department of Defense

The Pentagon awarded Perennial Autonomy a $500M-ceiling JIATF-401 contract for AI drone-on-drone counter-drone interceptors, making counter-drone autonomy a discrete DoD program of record.

Announced May 21, 2026 Merops, Bumblebee, Hornet Defence and emergency autonomy Source: DroneLife

Media & Entertainment 1 deployment

Telecom 1 deployment

SoftBank

SoftBank launched an OpenAI-powered 'Patching as a Service' providing vulnerability assessment, remediation planning and implementation advisory to the top 3,000 companies behind Japan's critical infrastructure, after internally validating the approach by running OpenAI's models across its own systems.

Announced Jun 16, 2026 Cybersecurity Source: SoftBank

Every entry names the organisation and links its source. Outcome figures are quoted as reported, never estimated. Vendor announcements without a named customer are excluded. Halted and reversed deployments are kept on purpose.