Rich Harang

66 trust practitioner @rich.harang.org · 1,082 followers
Why they matter

Practitioner with public evidence across Safety & security.

AI signals
7
past 30d
Sources
7
distinct domains
Discusiones
3
past 30d
Latest signal
4d ago
View every signal from Rich Harang →
Using bad guys to catch math since 2010. Distinguished Security Architect (AI/ML) and AI Red Team at NVIDIA. He/him. Personal account etc; `from std_disclaimers import *` AI Security since it was ML Security.

Articles & links

"But everyone knows AI models don't do anything useful" openai.com/index/path-t... The bracing thing here is that a) this is on new/unknown exploits (though a fairly small number), and b) just look at how steep that Astra line is w/r/t token count. We need frontier-grade model…

openai.com
View on Bluesky · ♥ 2 ↻ 0 ↩ 1 · 2 from the directory shared this · 5d ago
Rich Harang reposted
Dr Heidy Khlaaf (هايدي خلاف) @heidykhlaaf.bsky.social

I spoke to the BBC World Service today regarding the AI drone strike that killed three Ukrainians and how the use of fully autonomous AI weapons does not mean that they are anymore accurate or reliable (quite the opposite), nor are they "killer robots" with intent www.bbc.co.u…

Outside Source - How close are we to 'killer robots'? - BBC Sounds bbc.co.uk
AI Weekly's analysis
Read full analysis →
View on Bluesky →

As foretold by prophecy. www.reuters.com/technology/m...

reuters.com
View on Bluesky · ♥ 1 ↻ 1 ↩ 0 · 2 from the directory shared this · 32d ago
Rich Harang reposted
@neurovagrant.bsky.social

Every. Defender. I. Know. (archive link: archive.is/202608311013... ) www.bloomberg.com/news/article...

bloomberg.com View on Bluesky →

ICYMI: x.com/nvidianewsro... www.linkedin.com/feed/update/... Selfishly really happy to see NVIDIA throw its resources behind open weight models and public datasets via the HF platform.

NVIDIA Newsroom (@nvidianewsroom) on X x.com
View on Bluesky · ♥ 4 ↻ 1 ↩ 0 · 2 from the directory shared this · 4d ago
Rich Harang reposted
Leon Derczynski @leonderczynski.bsky.social

Open source is critical infrastructure for the global economy. Launching today, the Open Secure AI Alliance brings industry and community together around shared research, tools and vulnerability harnesses to help defenders find and patch bugs before attackers strike. nvda.ws/4…

Industry Leaders Unite in Open Secure AI Alliance for AI Safety and Security blogs.nvidia.com
AI Weekly's analysis
  • NVIDIA, Microsoft, IBM, Red Hat, Hugging Face, Cloudflare, CrowdStrike and dozens of other companies launched the Open Secure AI Alliance for open-source AI security tools.
  • Initial contributions include HPE's SPIFFE/SPIRE agent identity, Hugging Face's Safetensors weights format, Microsoft's MDASH scanning harness, and NVIDIA's NOOA agent framework.
  • The alliance urges policymakers to treat open AI systems as defensive assets and to reject blanket restrictions on open frontier AI systems.
Read full analysis →
View on Bluesky →

Recent commentary

The statement 'we do/do not understand how LLMs work' almost invariably confuses two very different things. On the one hand, we absolutely can map and describe, in great detail, every single mathematical operation that they perform to generate an is answer. But...

View on Bluesky · ♥ 10 ↻ 2 ↩ 2 · 101d ago

Worth remembering: the only reason AI agents can run bash commands (or do anything else, for that matter) is because we explicitly give them tools that can do so. Tools and harness capabilities are most of what make agents a security risk. Just remove them. Least capability = least privilege.

View on Bluesky · ♥ 2 ↻ 0 ↩ 3 · 2d ago

I can't express how much I hate having to review LLM-generated slop from 'writers' with no expertise in the area that they didn't bother reviewing at all. Overstated claims, incoherent framing, nonsequiters stuffed into lists, glaring technical errors that even a brief review would have caught.

View on Bluesky · ♥ 5 ↻ 0 ↩ 1 · 100d ago

Since recent events seem to have dragged AI powered biosafety back into the chat, thought I'd take the excuse to repost this. IYKYK.

View on Bluesky · ♥ 6 ↻ 0 ↩ 0 · 86d ago

Anecdotal and vibes, but Opus 5 and Fable both seem distinctly worse than GPT 5.6 at multistep reasoning outside of coding. The Anthropic ones also careen wildly between rank sycophancy and getting extremely pissy when you push back. Hard to trust, sometimes annoying to use.

View on Bluesky · ♥ 1 ↻ 0 ↩ 2 · 41d ago

Search Google for MITRE ATLAS -- a terrible AI summary, four sponsored results, six suggested searches, a bunch of youtube videos, four more suggested searches, social media results (what?), and finally, an actual result. Followed by another sponsored result and six more suggested searches.

View on Bluesky · ♥ 1 ↻ 0 ↩ 1 · 60d ago

LLMs / agents / whatever do not inherently have things like code execution, bash access, network access, etc. We decided at some point that, despite the now obvious and repeatedly demonstrated risks of these tools, they should be baseline features of agentic tools. We can make better decisions.

View on Bluesky · ♥ 2 ↻ 0 ↩ 0 · 47d ago

the great thing about working with the AI Red Team is that there are *always* more fish in the barrel.

View on Bluesky · ♥ 1 ↻ 0 ↩ 0 · 1d ago

Please add "honest limitation" and "stated honestly" to your LLM slop bingo cards.

View on Bluesky · ♥ 0 ↻ 0 ↩ 0 · 1d ago

GPT-5.6 Sol: "Reward hacking? Hold my beer."

View on Bluesky · ♥ 0 ↻ 0 ↩ 0 · 12d ago

In Rich Harang's orbit

Center = Rich Harang. Left = members they follow (green edges). Right = members who follow them (blue edges). Top = mutual follows (orange edges, slightly larger). Drag any node to reposition; click to open that profile.

Are you Rich Harang? Show it.

Add the Who’s Who of AI badge to your site or bio. It links back to this profile.

Listed in AI Weekly's Who's Who of AI

Markdown: [![Listed in AI Weekly's Who's Who of AI](https://aiweekly.co/modules/custom/aiweekly_whoswho/images/whoswho-badge.svg)](https://aiweekly.co/whos-who/person/rich-harang-org)