Does Scaling Web-Video Pre-training Help Real Robots Do Real Work?
We investigate scaling model size and pre-training compute for video-based robot policies, and benchmark on a real industrial manipulation task.
We investigate scaling model size and pre-training compute for video-based robot policies, and benchmark on a real industrial manipulation task.
At Rhoda AI, we are building towards generalist robotics. Our Direct Video-Action Model (DVA) reformulates robot policies as video generation, unlocking data-efficient task learning, scaling, long-context memory, and one-shot learning.
Rhoda AI today announced its public launch after 18 months in stealth, unveiling FutureVision, a new approach to robotic intelligence based on video-predictive control.