Bernhardsson Links LLM Limits to Short Feedback Loops
Erik Bernhardsson shares idea on why intelligent people and models may underperform.
Erik Bernhardsson posted a theory that highly intelligent people sometimes underperform because they focus on verifiable rewards with short feedback loops. He suggested the same pattern could explain limits in large language models. Researcher Shreya Shankar shared the post, commenting that it illustrates how grad students and academics engage with ideas on social media. The packet contains only these two posts and the generated headlines summarizing them.
I've had a theory about super smart people who aren't always successful, that it boils down to focusing too much on verifiable rewards with short feedback loops. It feels like there's a lesson that might carry over to LLMs here.