UALM: Unified Audio Language Model for Understanding, Generation and Reasoning
机构 * CMU(卡内基梅隆大学) ; NVIDIA(英伟达) ; UMD(马里兰大学)
高校专区
机构 * CMU(卡内基梅隆大学) ; NVIDIA(英伟达) ; UMD(马里兰大学)
机构 * Soka University of America(美国早稻田大学) ; University of Maryland, College Park(马里兰大学学院市分校)
Comments This paper contains 19 pages and 3 figures. To be presented at the 2nd Workshop on Aligning Reinforcement Learning Experimentalists and Theorists (ARLET 2025) at NeurIPS 2025