From Pixels to Tokens: A Systematic Study of Latent Action Supervision for Vision-Language-Action Models

ICMLOral2026

Authors
Yihan Lin, Haoyang Li, Yang Li, Haitao Shen, Yihan Zhao, Chao Shao, Jing Zhang
Venue
ICML 2026
Track
Oral

TL;DR

No summary has been collected for this paper yet. Read the abstract at the authoritative source below.

Read the paper

Topics

vision-language

← All ICML 2026 Oral papers · Browse the whole archive