Why trust an AI reviewer if you don't trust the agent it checks?
A user describes Codex and Claude Code's auto-approval as asking another agent for permission—and questions why that reviewer should be trusted.
TLDR
A user questions whether Codex and Claude Code's auto-approval offers reassurance if permission comes from another agent whose judgment you may not trust. In a follow-up, they describe studying this through a model of the user and reviewer agents' utility functions—what each seeks to maximize. In that model, the acting agent repeatedly proposes actions, while reviewers compare them with a baseline and vote to approve or deny based on perceived utility.
Combined views
750
1 Source, first seen 15d ago