DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization
DiFlowDubber:通过跨模态对齐与同步实现的离散流匹配自动化视频配音
机构 * FPT Software AI Center, Vietnam(FPT软件人工智能中心,越南) ; KAIST, South Korea(韩国科学技术院) ; University of Alabama at Birmingham, USA(阿拉巴马大学伯明翰分校)
AI总结 本文提出DiFlowDubber框架,通过离散流匹配和两阶段训练策略,解决视频配音中内容准确性、表达语气、高质量音频和精确唇同步的问题。
Comments Accepted at CVPR 2026 Findings