VANE: Reliable Test-Time Training for Vision-Language-Action Models via Future Visual Representation Prediction
作者:Hongjin Ji, Guoyang Xia, Luoyang Sun, Fangxiang Feng, Lei Ren · 单位:The Chinese University of Hong Kong, Shenzhen, Beijing University of Posts and Telecommunications, Institute of Automation, Chinese Academy of Sciences, University of Chinese Academy of Sciences, Li Auto Inc. · 会议/期刊:arXiv preprint · 方向:cs.RO · 发布日期:2026-08-11