Bridging the Language Gap: Synthetic Voice Diversity via Latent Mixup for Equitable Speech Recognition
弥合语言鸿沟:通过潜在混合实现合成语音多样性以实现公平的语音识别
机构 * University of California Los Angeles, Department of Statistics(加州大学洛杉矶分校统计学系)
AI总结 本文提出了一种通过潜在混合提升合成语音多样性的方法,以改善低资源语言的语音识别性能。
Comments Accepted at ICML 2025 Workshop on Machine Learning for Audio