Podcast Episodes

Back to Search
SkipVLA: Let the Planner Move, Let the VLA Manipulate

SkipVLA proposes a hybrid architecture combining a classical planner with a Vision-Language-Action model for manipulation tasks, achieving 2.5x faste…

6 days, 18 hours ago

Short Long
View Episode
ForeDrive: A World Model Is Only Useful If It Changes the Plan

ForeDrive introduces a planning-relevant latent world model that guides autonomous driving through foresight, enabling the agent to anticipate future…

6 days, 18 hours ago

Short Long
View Episode
DO AS I DO: Making Human Video Executable

A new algorithm reconstructs and retargets monocular RGB (or synthetically generated) videos to high-DoF robot hands, enabling scalable dexterous man…

1 week ago

Short Long
View Episode
GAM: Ground the Target Before You Learn the Action

GAM is a robot foundation model that conditions action chunk prediction on 3D grounded inputs — including language, 3D point clouds, and bounding box…

1 week ago

Short Long
View Episode
Tune Slowly, Control Quickly: Learning a Better Robot Navigation Stack

Autonomous navigation in complex, unstructured environments poses a significant challenge, with traditional planners lacking adaptability and end-to-…

1 week, 1 day ago

Short Long
View Episode
THAW-VLA: World-Model Features Without World-Model Latency

This work distills frozen world-model representations into compact VLAs via a single feature-alignment loss at training time, with the teacher cached…

1 week, 2 days ago

Short Long
View Episode
Joga: Why a Soccer Humanoid Needs to Control Its Gaze

Joga demonstrates agile humanoid soccer on a Unitree G1 using an actuated neck and residual models to close the sim-to-real gap for both perception a…

1 week, 2 days ago

Short Long
View Episode
AthenaZero: Why Dynamic Manipulation Starts With Lower Inertia

AthenaZero is a bimanual manipulator designed to minimize inertia without compromising control authority. By using quasidirect drive actuation and tr…

1 week, 3 days ago

Short Long
View Episode
FROA-Drive and the Hard Part of Targeted Policy Repair

Vision–Language–Action (VLA) models have shown strong potential for end-to-end autonomous driving, yet their post-training commonly relies on expensi…

1 week, 3 days ago

Short Long
View Episode
Teach the Robot Where to Act—Then Make It Fast

We present a framework for assistive robot manipulation that addresses two fundamental challenges: efficient adaptation of large-scale models for sce…

1 week, 4 days ago

Short Long
View Episode

Love PodBriefly?

If you like Podbriefly.com, please consider donating to support the ongoing development.

Support Us