作者:Zelong Sun, Jun Wang, Kaicheng Yang, Tiancheng Gu, Ziyong Feng, Zhiwu Lu · 单位:Glint Lab, (LVLM)-based retrievers are efficient and scalable, directly encoding raw multimodal, Meng et al., 2025). Although efficient and scalable, directly encoding the, query elaboration. · 会议/期刊:arXiv preprint · 方向:cs.CV · 发布日期:2026-08-07
点击后会加入生成队列