Guardrails removed spam, off-topic, unclear, or duplicate replies.
Ask a question below.
Published answers will appear here.
🚨 New paper: Some, but not all, AI companies make corporately-loyal models. xAI, DeepSeek, Anthropic, & OpenAI models all downplay company controversies. Google, Meta, & Alibaba models don't. The findings are clear, but we are pretty confused as to why... 🧵 @LennartFinke
We tested 7 hypotheses, one for each company. xAI, DeepSeek, Anthropic, and OpenAI models all clearly have a tendency to downplay company controversies. p<10^-5 (our permutation tests bottomed out). Google, Meta, and Alibaba models don't seem to do this, though!
xAI is the worst, followed by DeepSeek, Anthropic, and OpenAI.
See our preregistration for this study here. https://osf.io/twqps/overview
Guardrails removed spam, off-topic, unclear, or duplicate replies.
Ask a question below.
Published answers will appear here.