Zvi Questions 'Most Aligned Model' Meaning
Steven Adler retweets Zvi's question on alignment terminology.
TLDR
Steven Adler, an independent AI safety researcher and former OpenAI employee focused on dangerous capabilities evaluations, retweeted a post by @TheZvi. The post asks OpenAI staff what the term 'most aligned model' should be interpreted as meaning exactly. The original post addresses OpenAI folks specifically and expresses uncertainty about the phrase. No further details or responses appear in the evidence packet.
Combined views
25.6K
2 Sources, first seen 27d ago