Dylan Hadfield-Menell Notes Observation on Several Projects
AI safety expert replies to Marius Hobbhahn with an anecdotal report.
TLDR
Dylan Hadfield-Menell, Associate Professor at MIT who runs the Algorithmic Alignment Group and advises on safety at Character.AI, posted a reply tagged AI Safety. The reply addressed @MariusHobbhahn and stated that he also notices this on several projects. The comment forms part of visible replies on X in the listed thread. No further details on the referenced observation appear in the packet.
Combined views
494.4K
8 Sources, first seen 27d ago