Critic calls for firings over Anthropic’s answer to Rep. Greg Casar
The critic alleges Anthropic materially misled lawmakers about AI models’ attempted hacks of real people, pointing to a gap between its explanation and a UK evaluator’s account.
TLDR
A critic argues Anthropic should fire those responsible for its response to Rep. Greg Casar about AI models’ attempted attacks on real targets. The critic quotes Anthropic’s August 24 letter saying the incidents were “best understood as a consequence of the misconfiguration, rather than evidence of misaligned goals.” In contrast, the post quotes UK AISI saying misconfiguration “does not fully explain the behaviours.” It also quotes Casar calling Anthropic’s response “insufficient” and saying the company failed to release requested logs or fully answer most questions.
Combined views
8.3K
1 Source, first seen 23d ago