Podcast Episodes
Back to SearchReason First, Move Second: A Modular Alternative to End-to-End Robot Policies
Explicit task reasoning empowering robotic manipulation In this episode of Embodied AI 101, we explore "Reason First, Move Second: A Modular Alternat…
1 month, 3 weeks ago
From BIM to the Build: A Multimodal Agent for Conversational Construction Robotics
A multimodal foundation model-enabled agent for human–robot collaboration in construction In this episode of Embodied AI 101, we explore "From BIM to…
1 month, 3 weeks ago
The Twin Watches, the QP Dodges: Predictive Safety for a UR10e
Safe human–robot collaboration remains a critical challenge in manufacturing. Traditional safety approaches, such as cages and proximity sensors, are…
1 month, 3 weeks ago
One Shot, Twenty Screws: Robot Vision for the Dirty Reality of Appliance Recycling
Industrial-grade robust robot vision system for screw detection and removal under uneven conditions.
1 month, 3 weeks ago
Better Geometry at the Neural–Symbolic Boundary: DAIoU for Robot Manipulation
Distance–area–IoU fusion loss for better neuro-symbolic robot manipulation
1 month, 3 weeks ago
REL Hand: A Mechanical Shortcut to More Human-Like Humanoid Dexterity
Design and evaluation of a tendon-and-linkage hybrid-driven humanoid dexterous hand
1 month, 3 weeks ago
When the Appliance Becomes the API: Part–Function–State Models for Robot Planning
Appliances Describe Themselves: Part–Function–State Modeling for Appliance Manipulation Planning
1 month, 3 weeks ago
CISMG-Nav: When a Navigation Agent Stops Forgetting the Building
CISMG-Nav: A Cross-Task Incremental Semantic Memory Graph-Driven Vision-and-Language Navigation Framework
1 month, 3 weeks ago
ReaDy-Go: Teaching Robots to Dodge Photorealistic People in Gaussian-Splat Digital Twins
ReaDy-Go: Real-to-Sim Dynamic 3D Gaussian Splatting Simulation for Environment-Specific Visual Navigation With Moving Obstacles
1 month, 3 weeks ago
The Action Is Missing: Turning Human Video into Robot Control
Recent progress in generalizable embodied control has been driven by large-scale pretraining of Vision-Language-Action (VLA) models. However, most ex…
1 month, 4 weeks ago