ShadowDancer: Teaching Video World Models Any Action by Learning Unified Dynamics Representations from a Video and Its Shadow figure
AlphaXiv 中文概览(可滚动查看)