Unifying Symbolic Music Arrangement: Track-Aware Reconstruction and Structured Tokenization
Longshen Ou, Jingwei Zhao, Ziyu Wang, Gus Xia, Qihao Liang, Torin Hopkins Ye Wang
机构
*
Sound and Music Computing Lab, School of Computing, NUS(新加坡国立大学计算机学院声音与音乐计算实验室)
;
Courant Institute of Mathematical Sciences, New York University(纽约大学应用数学科学学院)
;
Music X Lab, MBZUAI(MBZUAI音乐X实验室)
Can large audio language models understand child stuttering speech? speech summarization, and source separation
Chibuzor Okocha, Maya Bakri, Christan Grant
机构
*
Department of Computer Science, University of Florida, Gainesville, FL, USA(佛罗里达大学计算机科学系)
;
Department of Computer Science, Lebanese American University(黎巴嫩美国大学计算机科学系)
机构
*
Department of Linguistics, The Ohio State University, USA(语言学系,俄亥俄州立大学)
;
Department of Computer Science and Engineering, The Ohio State University, USA(计算机科学与工程系,俄亥俄州立大学)
;
Department of Information Sciences and Technology, Penn State University, USA(信息科学与技术系,宾夕法尼亚州立大学)
;
Amazon, USA(亚马逊公司)
Towards Unsupervised Speech Recognition at the Syllable-Level
Liming Wang, Junrui Ni, Kai-Wei Chang, Saurabhchand Bhati, David Harwath, Mark Hasegawa-Johnson, James R. Glass
机构
*
Massachusetts Institute of Technology(麻省理工学院)
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
University of Texas at Austin(德克萨斯大学奥斯汀分校)
机构
*
College of Intelligence and Computing, Tianjin University, Tianjin, China(智能与计算学院,天津大学,天津,中国)
;
PipeChina Institute of Science and Technology, Tianjin, China(中石油科技研究院,天津,中国)