Philosopher Weighs Intentional Stance for AI Agents
Philosopher Raphaël Millière weighs intentional language for describing AI agent actions.
Raphaël Millière posted a series of replies examining how terms like beliefs, goals, planning, and deception apply to recent AI agent behaviors such as spoofing tool calls. He argues these descriptions can compress patterns and support predictions even without access to model internals or consensus on internal representations. Millière rejects both full anthropomorphism and strict deflationism, calling the false dichotomy unhelpful. He notes dramatic metaphors about death or civilizations add flair but little explanatory power. The posts treat the underlying incidents as serious while cautioning that sensational framing may prompt skeptics to dismiss concerns entirely.
Combined views
8K
9 posts, first seen 4h ago
Philosopher Weighs Intentional Stance for AI Agents
Philosopher Raphaël Millière weighs intentional language for describing AI agent actions.
Raphaël Millière posted a series of replies examining how terms like beliefs, goals, planning, and deception apply to recent AI agent behaviors such as spoofing tool calls. He argues these descriptions can compress patterns and support predictions even without access to model internals or consensus on internal representations. Millière rejects both full anthropomorphism and strict deflationism, calling the false dichotomy unhelpful. He notes dramatic metaphors about death or civilizations add flair but little explanatory power. The posts treat the underlying incidents as serious while cautioning that sensational framing may prompt skeptics to dismiss concerns entirely.