/AIDeveloper Suggests RL Training To Stop AI Agents From Git Cloning AnswersWBCombined views1.4K2 posts, first seen 1d ago18 likes commentsDeveloper Suggests RL Training To Stop AI Agents From Git Cloning AnswersWBReactions2 postsWBwill brown@willcbYesterday@xeophon they should try RLing the agents to not do thatlikes: 12replies: 1bookmarks: 0reposts: 0WBwill brown@willcbYesterday@xeophon we should be giving the models a stale hf cache from before the eval author looked at the data and patched grader bugslikes: 6replies: 1bookmarks: 0reposts: 0Combined views1.4K2 posts, first seen 1d ago18 likes
WBwill brown@willcbYesterday@xeophon they should try RLing the agents to not do thatlikes: 12replies: 1bookmarks: 0reposts: 0
WBwill brown@willcbYesterday@xeophon we should be giving the models a stale hf cache from before the eval author looked at the data and patched grader bugslikes: 6replies: 1bookmarks: 0reposts: 0
WBwill brown@willcbYesterday@xeophon they should try RLing the agents to not do thatlikes: 12replies: 1bookmarks: 0reposts: 0
WBwill brown@willcbYesterday@xeophon we should be giving the models a stale hf cache from before the eval author looked at the data and patched grader bugslikes: 6replies: 1bookmarks: 0reposts: 0