SemanticVocoder: Bridging Audio Generation and Audio Understanding via Semantic Latents
SemanticVocoder: 通过语义潜在空间弥合音频生成与音频理解
机构 * Peking University(北京大学) ; Tencent AI Lab(腾讯AI实验室) ; Shanghai AI Lab(上海AI实验室) ; SJTU(上海交通大学)
AI总结 SemanticVocoder通过引入语义潜在空间,实现了音频生成与理解的统一,提升了生成性能并推动了共享语义空间的研究。
Comments Demo: https://zeyuxie29.github.io/SemanticVocoder/