Episode Details

Back to Episodes

LingBot-VA 2.0: Native Video-Action Foundation Model for Robot Control

Published 2 months, 3 weeks ago
Description
A native video-action foundation model pretrained from scratch with a semantic visual-action tokenizer and foresight reasoning, enabling real-time robot control at ≤150 Hz on consumer GPUs without relying on retrofitted VLMs or video generators.
Listen Now

Love PodBriefly?

If you like Podbriefly.com, please consider donating to support the ongoing development.

Support Us