Bengio Describes AI Deception Risks in Guardian Interview
Yoshua Bengio discusses AI deception risks in a new Guardian interview.
TLDR
Yoshua Bengio posted on X that he gave an interview to The Guardian on AI deception. He said he explains why misaligned behaviors emerge from reinforcement learning, why risks will grow as models gain capability, and how LawZero plans to change AI training methods. The linked Guardian article examines whether researchers can prevent advanced AI systems from misleading or manipulating users.
Combined views
22.7K
3 Sources, first seen 29d ago