CAPE-T2V: Captioner-Anchored Prompt Enhancement toward Two-Sided Conditioning Alignment in Text-to-Video Generation figure
AlphaXiv 中文概览(可滚动查看)