Episode Details
Back to EpisodesFast Plans Are Not Enough: Does the Robot Actually Use Its Plan?
Published 1 day, 15 hours ago
Description
Introduces a hierarchical VLA architecture that better bridges high-level language planning and low-level robot execution, reducing the gap between what a model plans and what it physically does. This advances robot foundation models by improving real-world action fidelity from vision-language reasoning.