GRASP: Granularity-Aware Region Alignment and Semantic Prototype Learning for Fine-Grained Cross-Modal Understanding in Drone Views
作者:Jiahui Cui, Yan Zhao, Kan Wei, Enze Zhu, Peirong Zhang, Lei Wang, Yiru Wang · 单位:Aerospace Information Research Institute, Chinese Academy of Sciences, Key Laboratory of Target Cognition and Application Technology (TCAT), University of Chinese Academy of Sciences, School of Electronic, Electrical and Communication Engineering, University of Chinese Academy of Sciences, Aerospace Information Research Institute, Chinese Academy of Sciences Beijing China, Key Laboratory of Target Cognition and Application Technology (TCAT) Beijing China, University of Chinese Academy of Sciences Beijing China, School of Electronic, Electrical and Communication Engineering, University of Chinese Academy of Sciences Beijing China · 会议/期刊:arXiv preprint · 方向:cs.CV · 发布日期:2026-08-11