Episode Details
Back to EpisodesOne Model to See, Plan, and Act: Introducing EO-1 for Embodied AI
Published 2 months, 1 week ago
Description
A 3B parameter unified decoder-only transformer that interleaves vision, text, and action tokens for perception, planning, reasoning, and control in a single model. Trained on the 1.5M-sample EO-1.5M dataset with strong results across manipulation tasks and benchmarks.