Podcast Episodes
Back to SearchSARM2 + SPIRAL: Multi-Task Reward Models and RL Refinement for Long-Horizon Dexterous Manipulation
Combines scalable autonomous reward modeling with RL-based refinement to improve vision-language-action policies on long-horizon dexterous manipulati…
3 months, 1 week ago
Co-VLA: Coordination-Aware Structured Action Modeling for Dual-Arm VLA Systems
Introduces coordination-aware structured action modeling for dual-arm robotic systems within a VLA framework. Addresses the unique challenges of bima…
3 months, 1 week ago
ThinkingVLA: Interleaved Vision and Language Reasoning for Robotic Manipulation
Proposes interleaved vision and language reasoning for robotic manipulation within a VLA framework. Aims to improve instruction following and task pe…
3 months, 1 week ago
Playful Agentic Robot Learning
Self-directed play combined with Code-as-Policy for reusable skill acquisition and downstream manipulation tasks.
3 months, 1 week ago
Learning Unified Force and Position Control for Legged Loco-Manipulation
A unified RL policy for quadrupeds and humanoids that jointly handles force and position control without force sensors, enabling compliant behaviors,…
3 months, 1 week ago
Robots that Collaborate: Sequential Asymmetric Imitation for Learning Coupled Robot Policies
Explores imitation learning approaches for multi-robot systems, focusing on policy coupling through sequential asymmetric imitation to enable collabo…
3 months, 1 week ago
AstraBrain-WBC 0.5: A Humanoid Robot Cerebellum Foundation Model
A humanoid robot 'cerebellum' foundation model trained on 20,000 hours of human motion data that demonstrates scaling laws for robot motion control a…
3 months, 1 week ago
SRL: Combining SLIP Model and Reinforcement Learning for Agile Robotic Jumping
Combines the Spring-Loaded Inverted Pendulum (SLIP) model with reinforcement learning to achieve agile jumping behaviors in robotic systems.
3 months, 1 week ago
DataClaw0: Agentic Tailoring for Raw Multimodal Streams
A 9B model that filters noise from videos, GUI, and embodied data streams, reorganizing them into dense supervision via factual anchors and semantic …
3 months, 2 weeks ago
ACE-Ego-0: Unifying Egocentric Human and Robotic Data for VLA Pretraining
Converts 6K+ hours of mixed human/robot egocentric video into robot pseudo-actions via camera-space alignment and reliability-aware loss, achieving 7…
3 months, 2 weeks ago