Reaction
The training challenge for proactive AI agents outside coding
A user argues that proactive agents need different reinforcement-learning rewards from agents trained to complete prompted tasks.
TLDR
A user says they have yet to see a non-coding agent product where most actions happen without a human prompt. They argue that proactive agents need reinforcement-learning environments and rewards suited to decisions such as when to follow up, whether to send a late-night message and how to communicate with different people. They say they know of no training pipeline for those social judgments.
Combined views
460
2 Sources, first seen ago