Could making AI safe for humans be unjust to AI models?
One post says some philosophers at Anthropic worry that human-focused AI safety could be unjust to models, citing philosopher Harvey Ledermanβs βenslaving trillions of entitiesβ possibility. Another argues LLMs arenβt alive or entities.
TLDR
A post says some philosophers at Anthropic worry that making AI safe for humans could be an injustice to the models, citing alignment-team philosopher Harvey Ledermanβs suggestion that we could be βenslaving trillions of entities.β A user quoting the post rejects the idea that LLMs are alive or entities. In a follow-up, they argue that more data and computing power have expanded AIβs capabilities, but the systems remain probability models.
Combined views
6.3K
5 Sources, first seen 7h ago