Ex-Omni-2D: Expressive Omni-Modal Dialogue Models with Native Visual Presence
作者:Haoyu Zhang, Zhipeng Li, Xiaoying Tang, Tianshu Yu, Yiwen Guo · 单位:The Chinese University of Hong Kong, Shenzhen, This factorization is motivated by data availability, systems explore speech-centered instruction follow- · 会议/期刊:arXiv preprint · 方向:cs.AI · 发布日期:2026-08-12