Beyond Visual Ambiguity: Guiding Robust Monocular Depth Estimation in Challenging Scenarios via Detailed Long Captions
作者:Junrui Zhang, Jiaqi Li, Yiran Wang, Liao Shen, Zhiguo Cao · 单位:School of Artificial Intelligence and Automation, Huazhong University of Science and Technology, School of Artificial Intelligence and Automation, Huazhong University of Science and Technology Wuhan China · 会议/期刊:arXiv preprint · 方向:cs.CV · 发布日期:2026-07-31