Invisible Shortcuts: Why Vision Encoders Know Your Camera
作者:Vladan Stojnić, Ryan Ramos, Giorgos Kordopatis-Zilos, Noa Garcia, Giorgos Tolias · 单位:VRG, FEE, Czech Technical University in Prague, The University of Osaka, sion, whether through categorical labels (ImageNet) or billion-scale cap-, vision, including object–background dependencies [21,45,64], color–label correla-, similar conditions, creating systematic associations between labels or captions, vision: categorical label supervision on ImageNet and caption supervision on · 会议/期刊:arXiv preprint · 方向:cs.CV · 发布日期:2026-08-07