A consequence of how frontier models are trained is motivated reasoning, a phenomenon well studied in humans and discussed in this podcast from Palisade Research.
Original headline:AI Hacking Incidents with Tim Hua
Track only the AI that matters to you
Your own agent, watching your companies and topics.
Build your agent →
We use essential cookies to keep the site working (login, form security). With your permission, we also use analytics cookies to understand how you use the site.
Privacy policy