Podcast Episodes

Back to Search
SARM2 + SPIRAL: Multi-Task Reward Models and RL Refinement for Long-Horizon Dexterous Manipulation

Combines scalable autonomous reward modeling with RL-based refinement to improve vision-language-action policies on long-horizon dexterous manipulati…

3 months, 1 week ago

Short Long
View Episode
Co-VLA: Coordination-Aware Structured Action Modeling for Dual-Arm VLA Systems

Introduces coordination-aware structured action modeling for dual-arm robotic systems within a VLA framework. Addresses the unique challenges of bima…

3 months, 1 week ago

Short Long
View Episode
ThinkingVLA: Interleaved Vision and Language Reasoning for Robotic Manipulation

Proposes interleaved vision and language reasoning for robotic manipulation within a VLA framework. Aims to improve instruction following and task pe…

3 months, 1 week ago

Short Long
View Episode
Playful Agentic Robot Learning

Self-directed play combined with Code-as-Policy for reusable skill acquisition and downstream manipulation tasks.

3 months, 1 week ago

Short Long
View Episode
Learning Unified Force and Position Control for Legged Loco-Manipulation

A unified RL policy for quadrupeds and humanoids that jointly handles force and position control without force sensors, enabling compliant behaviors,…

3 months, 1 week ago

Short Long
View Episode
Robots that Collaborate: Sequential Asymmetric Imitation for Learning Coupled Robot Policies

Explores imitation learning approaches for multi-robot systems, focusing on policy coupling through sequential asymmetric imitation to enable collabo…

3 months, 1 week ago

Short Long
View Episode
AstraBrain-WBC 0.5: A Humanoid Robot Cerebellum Foundation Model

A humanoid robot 'cerebellum' foundation model trained on 20,000 hours of human motion data that demonstrates scaling laws for robot motion control a…

3 months, 1 week ago

Short Long
View Episode
SRL: Combining SLIP Model and Reinforcement Learning for Agile Robotic Jumping

Combines the Spring-Loaded Inverted Pendulum (SLIP) model with reinforcement learning to achieve agile jumping behaviors in robotic systems.

3 months, 1 week ago

Short Long
View Episode
DataClaw0: Agentic Tailoring for Raw Multimodal Streams

A 9B model that filters noise from videos, GUI, and embodied data streams, reorganizing them into dense supervision via factual anchors and semantic …

3 months, 2 weeks ago

Short Long
View Episode
ACE-Ego-0: Unifying Egocentric Human and Robotic Data for VLA Pretraining

Converts 6K+ hours of mixed human/robot egocentric video into robot pseudo-actions via camera-space alignment and reliability-aware loss, achieving 7…

3 months, 2 weeks ago

Short Long
View Episode

Love PodBriefly?

If you like Podbriefly.com, please consider donating to support the ongoing development.

Support Us