‘If you build something vastly smarter than you, it better be on your side’: can we stop AI from deceiving us?
2 directory members surfaced this signal.
“In an interview in this article for @theguardian.com about AI deception, I explain why misaligned behaviors emerge from reinforcement learning, why they will continue to pose risks as models become more capable, and how we intend to rethink how we train AI …” evidence ↗
Concern & critique
2 expertsRisks, limits and unintended consequences.
“In an interview in this article for @theguardian.com about AI deception, I explain why misaligned behaviors emerge from reinforcement learning, why they will continue to pose risks as models become more capable, and how we intend to rethink how we train AI …”