Rich Harang

66 trust practitioner @rich.harang.org · 1,082 followers
Why they matter

Practitioner with public evidence across Safety & security.

AI signals
3
past 30d
Sources
3
distinct domains
Discussões
5
past 30d
Latest signal
1d ago
View every signal from Rich Harang →
Using bad guys to catch math since 2010. Distinguished Security Architect (AI/ML) and AI Red Team at NVIDIA. He/him. Personal account etc; `from std_disclaimers import *` AI Security since it was ML Security.

Articles & links

Rich Harang reposted
Leon Derczynski @leonderczynski.bsky.social

Open source is critical infrastructure for the global economy. Launching today, the Open Secure AI Alliance brings industry and community together around shared research, tools and vulnerability harnesses to help defenders find and patch bugs before attackers strike. nvda.ws/4…

Industry Leaders Unite in Open Secure AI Alliance for AI Safety and Security blogs.nvidia.com
AI Weekly's analysis
  • NVIDIA, Microsoft, IBM, Red Hat, Hugging Face, Cloudflare, CrowdStrike and dozens of other companies launched the Open Secure AI Alliance for open-source AI security tools.
  • Initial contributions include HPE's SPIFFE/SPIRE agent identity, Hugging Face's Safetensors weights format, Microsoft's MDASH scanning harness, and NVIDIA's NOOA agent framework.
  • The alliance urges policymakers to treat open AI systems as defensive assets and to reject blanket restrictions on open frontier AI systems.
Read full analysis →
View on Bluesky →

Recent commentary

The statement 'we do/do not understand how LLMs work' almost invariably confuses two very different things. On the one hand, we absolutely can map and describe, in great detail, every single mathematical operation that they perform to generate an is answer. But...

View on Bluesky · ♥ 10 ↻ 2 ↩ 2 · 60d ago

I can't express how much I hate having to review LLM-generated slop from 'writers' with no expertise in the area that they didn't bother reviewing at all. Overstated claims, incoherent framing, nonsequiters stuffed into lists, glaring technical errors that even a brief review would have caught.

View on Bluesky · ♥ 5 ↻ 0 ↩ 1 · 60d ago

Since recent events seem to have dragged AI powered biosafety back into the chat, thought I'd take the excuse to repost this. IYKYK.

View on Bluesky · ♥ 6 ↻ 0 ↩ 0 · 45d ago

Anecdotal and vibes, but Opus 5 and Fable both seem distinctly worse than GPT 5.6 at multistep reasoning outside of coding. The Anthropic ones also careen wildly between rank sycophancy and getting extremely pissy when you push back. Hard to trust, sometimes annoying to use.

View on Bluesky · ♥ 1 ↻ 0 ↩ 2 · 22h ago

Search Google for MITRE ATLAS -- a terrible AI summary, four sponsored results, six suggested searches, a bunch of youtube videos, four more suggested searches, social media results (what?), and finally, an actual result. Followed by another sponsored result and six more suggested searches.

View on Bluesky · ♥ 1 ↻ 0 ↩ 1 · 20d ago

LLMs / agents / whatever do not inherently have things like code execution, bash access, network access, etc. We decided at some point that, despite the now obvious and repeatedly demonstrated risks of these tools, they should be baseline features of agentic tools. We can make better decisions.

View on Bluesky · ♥ 2 ↻ 0 ↩ 0 · 6d ago

In Rich Harang's orbit

Center = Rich Harang. Left = members they follow (green edges). Right = members who follow them (blue edges). Top = mutual follows (orange edges, slightly larger). Drag any node to reposition; click to open that profile.