Episode Details
Back to EpisodesMolmoBot: Training Robot Manipulation Entirely in Simulation
Published 6 months, 1 week ago
Description
Vision-language-action (VLA) model enabling zero-shot sim-to-real transfer for mobile manipulation tasks, trained entirely in simulation without real robot data, achieving 79.2% success on real-world DROID benchmarks outperforming baselines by 2x.