SapiensID 2.0: Aligning Human Recognition Foundation Models with Human Perception
作者:Yiyang Su, Jie Zhu, Feng Liu, Anil K. Jain, Xiaoming Liu · 单位:Michigan State University, Drexel University, University of North Carolina, Chapel Hill, face similarities, which do not provide semantic labels. On the temporal front, WebBody consists of, more than 90% static images. Compiling and labeling a video dataset of comparable scale to learn, Simultaneously, to model kinematic continuity without requiring a massive, newly labeled video · 会议/期刊:arXiv preprint · 方向:cs.CV · 发布日期:2026-08-12