Greenblatt Discusses Reward-Seeking AI Takeover Risks
Greenblatt shares views on rapid takeoff and misalignment threats from reward-seeking systems.
TLDR
Dwarkesh Patel hosted Ryan Greenblatt on his podcast to debate recursive self-improvement after human-level AI. Greenblatt's median timeline for full automation of AI R&D is late 2030 or early 2031, with a modal guess of mid 2029. The talk examined whether progress could compress years of advances into one year and focused on takeover risks from reward-seeking AIs. Greenblatt noted that calls on large experiments may remain a bottleneck. Patel questioned whether misalignment would resemble minor human conflicts rather than catastrophic outcomes. Colleague Alex Mallen has written in detail on these threat models.
Combined views
603.2K
12 Sources, first seen 50d ago