Reaction
The case against scaling AI past a certain point before solving mechanistic interpretability
A user argues that mechanistic interpretability is a minimum requirement for making AI alignment an engineering discipline.
TLDR
A user argues that AI alignment cannot become an engineering discipline without solving mechanistic interpretability. In a reply to their own post, they call scaling AI past a certain point without it “suicidal,” say such scaling should not be allowed anywhere, and add that international cooperation would be hard but possible.
Combined views
2.3K
1 Source, first seen ago