The science and politics of aligning general-purpose AI
A post argues that aligning general-purpose AI hinges on political choices about desirable values, not just model training.
TLDR
The writer argues that training a model to follow data is relatively straightforward, but aligning general-purpose AI is harder: it requires deciding which concepts reflect desirable values and verifying that they constrain other concepts. The post calls those decisions philosophical and political, and claims frontier labs have incentives to delay announcing a full technical solution.
Combined views
263
2 Sources, first seen ago