机构
*
SuZhou Automotive Research Institute of Tsinghua University(清华大学苏州汽车研究院)
;
Department of Electrical and Electronic Engineering, The University of Hong Kong(香港大学电子与电气工程系)
;
Hyundai Motor Advanced Tech. R&D Center
;
School of Vehicle and Mobility, Tsinghua University(清华大学车辆与移动系统学院)
ManchuTTS: Towards High-Quality Manchu Speech Synthesis via Flow Matching and Hierarchical Text Representation
曼juTTS: 通过流匹配和分层文本表示实现高质量曼ju语音合成
Suhua Wang, Zifan Wang, Xiaoxin Sun, D. J. Wang, Zhanbo Liu, Xin Li
机构
*
Department of Computer Science, Changchun Humanities and Sciences College(计算机科学系,长春人文科学学院)
;
School of Information Science and Technology, Northeast Normal University(信息科学与技术学院,东北师范大学)
;
College of Optical Science and Engineering, Zhejiang University(光学科学与工程学院,浙江大学)
Bowen Shi, Andros Tjandra, John Hoffman, Helin Wang, Yi-Chiao Wu, Luya Gao, Julius Richter, Matt Le, Apoorv Vyas, Sanyuan Chen, Christoph Feichtenhofer, Piotr Dollár, Wei-Ning Hsu, Ann Lee
Text-Queried Audio Source Separation via Hierarchical Modeling
通过分层建模实现文本查询的音频源分离
Xinlei Yin, Xiulian Peng, Xue Jiang, Zhiwei Xiong, Yan Lu
机构
*
University of Science and Technology of China(中国科学技术大学)
;
School of Information and Communication Engineering, Communication University of China(中国通信大学信息与通信工程学院)
;
Microsoft Research Asia(微软亚洲研究院)
机构
*
Fudan University(复旦大学)
;
Shanghai AI Laboratory(上海人工智能实验室)
;
The Chinese University of Hong Kong(香港中文大学)
;
Shanghai Innovation Institute(上海创新研究院)
;
Beihang University(北京航空航天大学)
;
Shanghai Jiao Tong University(上海交通大学)
Unifying Symbolic Music Arrangement: Track-Aware Reconstruction and Structured Tokenization
Longshen Ou, Jingwei Zhao, Ziyu Wang, Gus Xia, Qihao Liang, Torin Hopkins Ye Wang
机构
*
Sound and Music Computing Lab, School of Computing, NUS(新加坡国立大学计算机学院声音与音乐计算实验室)
;
Courant Institute of Mathematical Sciences, New York University(纽约大学应用数学科学学院)
;
Music X Lab, MBZUAI(MBZUAI音乐X实验室)