SAM3D-Guided Object-Centric Representation Alignment for Vision-Language-Action Models
作者:Zonghe Liu (1), Shanyuan Jie (2), Xiaoquan Sun (3), Chen Cao (1), Zetian Xu (1), Zongsheng Liu (4), Jiayu Chen (1 and 5) ((1) University of Hong Kong, (2) Shenzhen Institutes of Advanced Technology, Chinese Academy of Sciences, (3) Huazhong University of Science and Technology, (4) Beijing University of Aeronautics and Astronautics, (5) Infiforce) · 单位:University of Hong Kong, Shenzhen Institutes of Advanced Technology, Chinese Academy of Sciences, Huazhong University of Science and Technology, Beijing University of Aeronautics and Astronautics · 会议/期刊:arXiv preprint · 方向:cs.RO · 发布日期:2026-07-29