Podcast Episodes
Back to SearchSkipVLA: Let the Planner Move, Let the VLA Manipulate
SkipVLA proposes a hybrid architecture combining a classical planner with a Vision-Language-Action model for manipulation tasks, achieving 2.5x faste…
6 days, 18 hours ago
ForeDrive: A World Model Is Only Useful If It Changes the Plan
ForeDrive introduces a planning-relevant latent world model that guides autonomous driving through foresight, enabling the agent to anticipate future…
6 days, 18 hours ago
DO AS I DO: Making Human Video Executable
A new algorithm reconstructs and retargets monocular RGB (or synthetically generated) videos to high-DoF robot hands, enabling scalable dexterous man…
1 week ago
GAM: Ground the Target Before You Learn the Action
GAM is a robot foundation model that conditions action chunk prediction on 3D grounded inputs — including language, 3D point clouds, and bounding box…
1 week ago
Tune Slowly, Control Quickly: Learning a Better Robot Navigation Stack
Autonomous navigation in complex, unstructured environments poses a significant challenge, with traditional planners lacking adaptability and end-to-…
1 week, 1 day ago
THAW-VLA: World-Model Features Without World-Model Latency
This work distills frozen world-model representations into compact VLAs via a single feature-alignment loss at training time, with the teacher cached…
1 week, 2 days ago
Joga: Why a Soccer Humanoid Needs to Control Its Gaze
Joga demonstrates agile humanoid soccer on a Unitree G1 using an actuated neck and residual models to close the sim-to-real gap for both perception a…
1 week, 2 days ago
AthenaZero: Why Dynamic Manipulation Starts With Lower Inertia
AthenaZero is a bimanual manipulator designed to minimize inertia without compromising control authority. By using quasidirect drive actuation and tr…
1 week, 3 days ago
FROA-Drive and the Hard Part of Targeted Policy Repair
Vision–Language–Action (VLA) models have shown strong potential for end-to-end autonomous driving, yet their post-training commonly relies on expensi…
1 week, 3 days ago
Teach the Robot Where to Act—Then Make It Fast
We present a framework for assistive robot manipulation that addresses two fundamental challenges: efficient adaptation of large-scale models for sce…
1 week, 4 days ago