Podcast Episodes
Back to SearchRobometer and the Future of Robotic Reward Modeling
New framework for scalable robotic reward modeling using trajectory comparisons to train general-purpose reward models.
4 months ago
Qwen-VLA: A Generalist Vision–Language–Action Robot Model
A single generalist VLA built on Qwen3.5-4B + 1.15B DiT flow-matching action decoder that unifies manipulation, navigation, and trajectory prediction…
4 months, 1 week ago
EXPO-FT: Sample-Efficient Reinforcement Learning Fine-Tuning for Vision-Language-Action Models
Extends the EXPO method with real-world RL post-training for VLAs using image observations, action chunking, DAgger, and on-the-fly Q-value maximizat…
4 months, 1 week ago
RoboMeter: A Dense Video-Language Reward Model for Robots
RoboMeter unifies frame-level progress and trajectory-level preferences into a single reward network trained on over a million robot trajectories, ge…
4 months, 1 week ago
RoboMeter: Learning Dense Rewards from Successes and Failures
RoboMeter trains dense reward models from both successful and failed robot trajectories, solving a key gap in prior methods that only learn from expe…
4 months, 1 week ago
MobileGym: A Controllable, Parallel Sandbox for Mobile GUI Agents
Browser-hosted mobile environment with JSON state, deterministic judges, and 256 parallel rollouts. Reports +40.7 real-device points after GRPO train…
4 months, 1 week ago
ANY2ANY: Efficient Cross-Embodiment Transfer for Humanoid Whole-Body Tracking
Introduces a method to transfer a Unitree G1 foundation policy (Gear-Sonic) to LimX Oli/Luna humanoids using only 1% of the original compute/data. Ac…
4 months, 1 week ago
TriSplat: Feed-Forward 3D Reconstruction with Triangulated Meshes
Outputs physics-engine-compatible triangle meshes directly from sparse, unposed images without Gaussian splatting or post-processing.
4 months, 1 week ago
MIKASA-Robo-VLA: A Memory-Intensive Benchmark for Vision-Language-Action Robotics
Releases a benchmark suite for systematically evaluating memory in Vision-Language-Action policies on tabletop manipulation tasks.
4 months, 1 week ago
PointWorld: Scaling 3D World Models for In-The-Wild Robotic Manipulation
Introduces large-scale 3D world models pretrained on diverse real-world video to enable robust robotic manipulation policies that generalize beyond s…
4 months, 1 week ago