XR-1: Towards Versatile Vision-Language-Action Models via Learning Unified Vision-Motion Representations
ICMLOral2026
TL;DR
No summary has been collected for this paper yet. Read the abstract at the authoritative source below.
Read the paper
Topics
vision-language