Episode Details
Back to EpisodesVistaVLA: Geometry- and Semantic-Aware 3D Gaussian-Grounded VLA for Robotic Manipulation
Published 2 months, 2 weeks ago
Description
Presents a vision-language-action model grounded in 3D Gaussian representations that integrates geometric and semantic awareness for improved robotic manipulation.