In an interview in this article for The
@guardian about AI deception, I explain why misaligned behaviors emerge from reinforcement learning, why they will continue to pose risks as models become more capable, and how we intend to rethink how we train AI systems at
@LawZero_ .