Reka AI Labs releases Rho-1 research preview for multimodal AI and robot actions
Reka describes one model for creating and reasoning across media, with unresolved limits in video consistency and editing.
TLDR
Reka AI Labs released Rho-1, a 19-billion-parameter research preview it says handles text, images, video and robot actions in one network. The company describes steerable video and shared state across conversations, and says it trained the model on 320 H100 GPUs in about three months. Its announcement also lists structural drift, unreliable object grounding across video, brittle edits and a 672-by-384 video resolution cap.
Combined views
27.7K
15 Sources, first seen ago
