Episode Details
Back to EpisodesAnchor the Semantics, Align the Actions: A Better Way to Fine-Tune VLAs
Published 2 months ago
Description
Proposes a VLA fine-tuning technique that preserves pretrained VLM representations while aligning language to actions, raising real-robot generalization success from 28% to 54%.