The tension between AI automation and tuning models for human feedback
A post sharing a video frames the Jev creator’s critique of RLHF around a mismatch: wanting to automate everything while tuning models to optimize for human feedback.
TLDR
A post sharing a video describes the Jev creator’s argument that reinforcement learning from human feedback (RLHF) is “guaranteed to disappoint” because it tells people what they want to hear. The post contrasts the ambition to automate everything with models being tuned to optimize for human feedback.
Combined views
8.1K
2 Sources, first seen 8h ago
The tension between AI automation and tuning models for human feedback
A post sharing a video frames the Jev creator’s critique of RLHF around a mismatch: wanting to automate everything while tuning models to optimize for human feedback.