The 'no plan' critique of AI alignment after The Curve
An attendee argues labs are counting on current AI to solve harder alignment problems, without a convincing backup.
TLDR
Writing after The Curve's third conference, an attendee calls the apparent AI safety strategy “No Plan”: align today's models, then have them help solve alignment for more capable systems. The writer argues mistakes in automated alignment could compound and says proposals to “pace the frontier” remain ill-defined. They favor limits on frontier-training compute to buy time, while doubting that better evaluations and monitoring alone would be enough.
Combined views
17.6K
1 Source, first seen ago