using Alice to red-team and monitor production AI systems for jailbreaks, prompt injection, and agentic misuse
Accueil › AI Use-Case Library › AI in security operations: 22 real deployments
AI in security operations: 22 real deployments
Named deployments in security operations, grouped by industry.
Software & Tech 15 deployments
Implemented chain-of-thought monitoring on all RL training and evaluations involving tools for GPT-5.6 Sol-class and higher models and all inference on the Astra model
Reported: monitoring overhead at roughly 20% of the inference compute being monitored
Used Wiz Red Agent AI to audit Snowflake's snowflake-connector-net repository and discover a shell-injection vulnerability introduced by a GitHub Copilot Autofix patch
Reported: Red Agent identified an exploitable shell-injection flaw; unauthenticated attacker exfiltrated the Jira token for [email protected] within five days, granting read access to engineering, security compliance and bug bounty projects
Using GPT-5.6 Sol Ultrafast mode internally for incident-response log analysis
deployed GPT-5.6-Cyber through the Daybreak Red program for offensive security vulnerability research, discovering previously unknown Chrome V8 vulnerabilities
Reported: model discovered two previously unknown V8 vulnerabilities in Chrome that can be chained to corrupt memory and bypass the V8 heap sandbox, patched by Google under CVE-2026-15903; model responds to 95% of sensitive security queries versus 57.3% for predecessor GPT-5.5-Cyber
deployed autonomous AI research system (HTTP Terminator) to test 30,000 candidate desync vectors against thousands of authorized websites to identify novel HTTP desync vulnerabilities
Reported: identified roughly 700 vulnerable targets including banks, government infrastructure, security products, and an airport; generated new attack classes including dual-matching Content-Length pattern, dangling-byte technique, and shared-parser confusion; exposed an Apache Traffic Server zero-day
Google uses AI tools to discover, validate, triage and fix Chrome vulnerabilities and is piloting twice-weekly security releases to absorb the resulting volume.
Reported: Chrome versions 149 and 150 fixed 1,072 bugs between them.
XBOW used its autonomous offensive-security agent to test Microsoft Bing Images and disclose two unauthenticated command-injection flaws.
Reported: The agent found CVE-2026-32194 and CVE-2026-32191, both rated CVSS 9.8; Microsoft fixed both before disclosure.
Meta deployed a free Facebook identity badge that uses a facial-recognition selfie to distinguish real users from AI-generated impostors.
Searchlight Cyber used GPT-5.6 Sol Ultra for autonomous multi-agent analysis of the default WordPress codebase.
Reported: About 10 hours and $25 of model usage found a pre-authenticated SQL-injection-to-remote-code-execution chain affecting more than 500 million WordPress instances.
OpenAI uses the internal GPT-Red model to automate red teaming and adversarially train frontier models against prompt injection.
Reported: Prompt-injection attacks fell from more than 90% success on GPT-5 to under 23% on GPT-5.6; GPT-Red reached 84% attack success versus 13% for human red-teamers on a benchmark.
Microsoft is reorganizing its cybersecurity business around AI-assisted vulnerability discovery and automated threat response.
Reported: The restructuring replaced senior executives and cut hundreds of roles.
Microsoft uses the multi-model MDASH agentic scanning harness to discover Windows vulnerabilities faster.
Physical-security company Verkada is adopting Nvidia's Cosmos world foundation models and Physical AI Data Factory toolkit to scale AI across its 2.4 million connected devices.
Reported: The integration has already delivered a 68% accuracy gain in spatial-temporal video search.
Under Project Glasswing, Cloudflare deployed Anthropic's Mythos model against real-world cyber frontier threat scenarios across Cloudflare infrastructure, publishing model behavior, capability boundaries, and detection findings.
Government & Public Sector 3 deployments
Installed more than 3,000 AI-powered license plate reader cameras via grants to cities and counties since 2023 for vehicle surveillance
Reported: Texas Gov. Abbott ordered state agencies to pause funding amid privacy concerns and misuse reports; cities and counties canceling contracts; Lufkin officer facing 100 counts for surveilling 11 people
Agents wore Meta's AI smart glasses during immigration enforcement operations across six states; internal memo now bars employees from wearing the devices due to risk of capturing sensitive data
The Gold Eagle federal clearinghouse uses frontier AI to scan government and private-sector systems for vulnerabilities and prioritize patches.
Military 2 deployments
Used DeepSeek to generate exploit code for cyberattacks
The Pentagon awarded Perennial Autonomy a $500M-ceiling JIATF-401 contract for AI drone-on-drone counter-drone interceptors, making counter-drone autonomy a discrete DoD program of record.
Media & Entertainment 1 deployment
Using facial recognition to identify and bar attorneys pursuing litigation from entering the venue
Telecom 1 deployment
SoftBank launched an OpenAI-powered 'Patching as a Service' providing vulnerability assessment, remediation planning and implementation advisory to the top 3,000 companies behind Japan's critical infrastructure, after internally validating the approach by running OpenAI's models across its own systems.
Every entry names the organisation and links its source. Outcome figures are quoted as reported, never estimated. Vendor announcements without a named customer are excluded. Halted and reversed deployments are kept on purpose.