GRAM: Spatial general-purpose audio representation models for real-world applications
GRAM:用于真实世界应用的空间通用音频表示模型
机构 * Donders Institute, Radboud University, Nijmegen, The Netherlands(多纳茨研究所,拉德堡德大学,尼姆韦根,荷兰) ; Mortimer B Zuckerman Institute, Columbia University, New York, United States(莫蒂默·B·祖克erman研究所,哥伦比亚大学,纽约,美国)
AI总结 GRAM通过多通道掩码自编码器学习空间音频表示,提升真实世界声学环境中的音频任务性能。
Comments Revise with RealSELD