Towards Robust and Controllable Text-to-Motion via Masked Autoregressive Diffusion
面向鲁棒且可控的文本到运动的掩码自回归扩散
机构 * State Key Laboratory of Virtual Reality Technology and Systems(虚拟现实技术与系统国家重点实验室) ; Hangzhou Innovation Institute(杭州创新院) ; Beihang University(北京航空航天大学)
AI总结 MoMADiff通过结合掩码建模与扩散过程,实现鲁棒且可控的文本到运动生成,优于现有方法。
Comments Accepted by ACM MM 2025