Podcast Episodes
Back to SearchBeyond Looking Human: Bio-Functional Robot Design for Cross-Scale Dexterity and Tool Use
Dexterous hands and robotic manipulators as physical interfaces for diverse environments. Proposes bio-functional mimicry beyond structural biomimicr…
1 month, 2 weeks ago
LDA-1B and the Case for Never Throwing Robot Data Away
Scales a latent dynamics action model to 1B parameters by ingesting diverse embodied datasets into a unified training pipeline for generalist robot p…
1 month, 2 weeks ago
Flex-π: One Robot Policy, Three Futures, Fifty-Six Ways to Run It
Jointly predicts future RGB frames, 3D pointmaps, DINO semantics, and actions within a unified world-action model. Can be deployed flexibly as a VLA,…
1 month, 3 weeks ago
Imagined Motion, Grounded Geometry: How Rest2Art Makes Closed Objects Articulate
An ECCV 2026 paper that converts a single static closed-state image into a simulation-ready articulated asset using video diffusion for joint hypothe…
1 month, 3 weeks ago
ComBodied Agents: When the Human Becomes the World Model
Proposes a new agent architecture that unifies digital and physical (embodied) action spaces, supporting human-state trajectories with explicit handl…
1 month, 3 weeks ago
A Sphere Every Robot Hand Can Speak
Introduces a unified hand action space that enables seamless cross-embodiment transfer for dexterous manipulation policies across different robot han…
1 month, 3 weeks ago
The Soft Gripper Pareto Frontier: Adaptability Without Surrendering Payload
Soft gripper systems have inherent flexibility and remarkable interactivity security. However, balancing load-bearing capacity, deformation, complexi…
1 month, 3 weeks ago
Dyna-2's Million-Hour Bet: When Human Video Starts Improving Robots
Pre-trained on 1 million hours of human video, Dyna-2 demonstrates scaling laws for world-action models across four orders of magnitude on human data…
1 month, 3 weeks ago
Think Deeply Only When It Matters: DySL-VLA and Action-Aware Dynamic Depth
Vision-Language-Action (VLA) models have shown remarkable success in robotic tasks like manipulation by fusing a language model's reasoning with a vi…
1 month, 3 weeks ago
The 100,000-Hour Bet: How Xiaomi-Robotics-1 Tries to Make Robot Scaling Real
We present Xiaomi-Robotics-1, a foundational vision-language-action (VLA) model capable of (1) following diverse language instructions to perform a w…
1 month, 3 weeks ago