Dr Heidy Khlaaf (هايدي خلاف)
Researcher with public evidence across AI research, Models & releases.
- AI signals
- 7 past 30d
- Sources
- 6 distinct domains
- Discusiones
- 1 past 30d
- Latest signal
- 20h ago
Articles & links
As someone who holds degrees in both CompSci and Philosophy, it's a relief to finally see a notable Philosopher break ranks with those who join tech companies to give them an heir of "intellectual credibility" while "pre-arming AI companies against criticism." www.ft.com/conte…
Very few people understand existing limitations of sandboxing, in addition to the tools and scaffolding AI are given to achieve this task: embroidery.io/blog/in-sand...
New! We hijack Claude Code(Sonnet 4.6,5/Opus 4.8) & Codex(GPT5.5) to achieve RCE when used to defensively assess an open-source/third-party library w/ prompt injections disseminated across its codebase. All without any skills, JSON, MCP, or config files required. ainowinstitut…
There will be a launch event happening today at 4PM BST/11 AM ET. You can sign up for it here: www.eventbrite.co.uk/e/launch-ana...
We’ve also summarized our findings and their policy implications given recent initiatives that seek to accelerate the use of defensive AI in Natsec and safety-critical infrastructure without consideration of these unmitigated risks: ainowinstitute.org/publications...
Recent commentary
The coverage on this OpenAI incident is abysmal. Use of the terms "rogue" and "loss of human control" lead to groupthink as people lack the critical skills to understand the difference between "autonomy" and faulty reward functions in AI on a task it was directed and given access to do.
In Dr Heidy Khlaaf (هايدي خلاف)'s orbit
Center = Dr Heidy Khlaaf (هايدي خلاف). Left = members they follow (green edges). Right = members who follow them (blue edges). Top = mutual follows (orange edges, slightly larger). Drag any node to reposition; click to open that profile.