Séb Krier Urges Broader AI Alignment Beyond Individual Models
The DeepMind policy lead published a post arguing alignment must cover systems and institutions in multi-agent settings.
Séb Krier, AGI Policy Development Lead at Google DeepMind, posted that alignment extends beyond model properties to the systems, institutions, protocols, and mechanism design governing multi-agent environments. His essay, titled Of Swarms and Sand Gods and published on the Cosmos Institute blog, states that better harnesses and protocols will be critical for safety. Ryan Lowe replied that the view has moved from rare to commonplace within a year. Replies on X largely praised the framing.
New post! Alignment doesn't just relate to a model's properties, but also the systems and institutions under which it operates. As we enter an increasingly multi-agent world, better harnesses, protocols, and mechanism design will be critical for safety. 🐜


