Podcast Episodes
Back to SearchWhen the Command Disappears: What Counts as Robot Intelligence?
Robot Independent Intelligence: When Language Vanishes in Vision-Language-Action Models
14 hours ago
Fast Plans Are Not Enough: Does the Robot Actually Use Its Plan?
Introduces a hierarchical VLA architecture that better bridges high-level language planning and low-level robot execution, reducing the gap between w…
1 day, 14 hours ago
Atlas's New Hands: Built to Learn, Not Just to Look Human
Atlas is upgraded with 13 DoF dexterous hands featuring tactile sensing, direct drive actuation, and sim-to-real RL training, designed for real tool …
1 day, 14 hours ago
Astra Gets a Body: What Zero-Shot Humanoid Control Actually Transfers
Astra demonstrates zero-shot open-ended pick-and-place on a physical humanoid robot, bridging simulation tooling directly to real-world execution wit…
2 days, 14 hours ago
Physical Coding: Robots That Check Their Work—and Keep the Fixes
A new paper introduces self-evolving coding agents for robotics that represent task state and execution as code, enabling robots to verify, recover f…
3 days, 14 hours ago
SpatialClaw: Give the VLM a Geometry Notebook
SpatialClaw is an NVIDIA spatial reasoning agent that uses Python-based code generation to perform 3D spatial reasoning, achieving +13.6 average impr…
3 days, 14 hours ago
Real-Time EXPO-FT: Slow VLA, Fast Corrections, Better Control
Async RL fine-tuning via EXPO-FT enables π0.5 to handle highly dynamic manipulation tasks such as ball balancing and striking in real time. The appro…
4 days, 14 hours ago
Aether: From Plausible Futures to Feasible Robot Plans
Aether's unified world model jointly learns environment dynamics and long-term physical state evolution under physical constraints, enabling robots t…
4 days, 14 hours ago
ARLI: Fixing the Missing State in Asynchronous Robot RL
Proposes a framework that restores the Markov property for RL fine-tuning of VLAs despite asynchronous inference delays, enabling effective RL traini…
5 days, 14 hours ago
Flex-π: Teach Geometry, Choose How Much to Imagine
A 6B-parameter world-action model that jointly predicts 3D pointmaps and DINO features alongside RGB frames, improving demonstration efficiency, gene…
5 days, 14 hours ago