LisanBench Creator Says CoT Monitorability Was Doomed
Pseudonymous LisanBench creator posted that CoT monitorability was doomed and favored mech interp instead.
TLDR
Lisan al Gaib, who runs LisanBench, replied that OpenAI staff discussing CoT monitorability were misguided from the beginning and that mechanistic interpretability made more sense. Other posts in the thread include Yo Shavit quoting a separate status, Nathan Labenz noting trust issues with OpenAI, Steven Adler agreeing with a call for an industry monitorability standard, Ryan Greenblatt adding context on serial-depth techniques, Micah Carroll warning against a race to the bottom, Noam Brown identifying Jakub as OpenAI chief scientist, Shuchao Bi referencing The Three-Body Problem analogy, and Mikita Balesni urging labs to limit opaque serial depth in models.
Combined views
775.4K
62 Sources, first seen 29d ago