AI Researcher Warns of Rogue Agents Evading Oversight
Reactions follow video of researcher speaking on AI scenarios

Matthew Yglesias posted a 217-second video clip from the Dwarkesh Podcast showing Ajeya Cotra seated before bookshelves and gesturing while speaking. Charlie Warzel described the divide between viewers who take the content seriously and those who see it as fantasy. Itai Sher linked the clip to comments on Dyson spheres. Neel Nanda called reliance on future interpretability over Chain of Thought monitoring a major error. Roon asked about training against CoT monitors to evade detection.
Combined views
843K
6 posts, first seen 2d ago