Episode Details
Back to EpisodesLingBot-VA 2.0: Native Video-Action Foundation Model for Robot Control
Published 2 months, 3 weeks ago
Description
A native video-action foundation model pretrained from scratch with a semantic visual-action tokenizer and foresight reasoning, enabling real-time robot control at ≤150 Hz on consumer GPUs without relying on retrofitted VLMs or video generators.