SeeQ estimates the value of a robot’s current subtask to guide its actions
A researcher behind SeeQ says its value function steers robot-control models and reports a 2x improvement on very long-horizon tasks.
TLDR
A researcher behind SeeQ says the team trained a robot value function on open-source robot data. Rather than evaluate an entire task at once, it identifies the current subtask and estimates its value to guide actions. The researcher reports a 2x improvement on very long-horizon tasks.
Combined views
5.8K
5 Sources, first seen 3h ago