机构
*
National University of Singapore(新加坡国立大学)
;
University of Toronto(多伦多大学)
;
Peking University(北京大学)
;
Sichuan University(四川大学)
;
Zhejiang University(浙江大学)
Vision Language Models Map Logos to Text via Semantic Entanglement in the Visual Projector
Sifan Li, Hongkai Chen, Yujun Cai, Qingwen Ye, Liyang Chen, Junsong Yuan, Yiwei Wang
机构
*
University of California, Merced(加州大学梅尔德分校)
;
vivo Mobile Communication Co., Ltd.(vivo移动通信有限公司)
;
University of Queensland(昆士兰大学)
;
UCLA(加州大学洛杉矶分校)
;
University at Buffalo(布法罗大学)
Unlocking LLM Safeguards for Low-Resource Languages via Reasoning and Alignment with Minimal Training Data
Zhuowei Chen, Bowei Zhang, Nankai Lin, Tian Hou, Lianxi Wang
机构
*
Guangdong University of Foreign Studies(广东外语外贸大学)
;
Guangzhou Key Laboratory of Multilingual Intelligent Processing(广州多语智能处理重点实验室)
;
University of Pittsburgh(匹兹堡大学)
机构
*
Guilin University of Electronic Technology(桂林电子科技大学)
;
The Ohio State University(俄亥俄州立大学)
;
University of Wisconsin-Madison(威斯康星大学麦迪逊分校)
;
University of California, Berkeley(加州大学伯克利分校)
Humanoid Agent via Embodied Chain-of-Action Reasoning with Multimodal Foundation Models for Zero-Shot Loco-Manipulation
Congcong Wen, Geeta Chandra Raju Bethala, Yu Hao, Niraj Pudasaini, Hao Huang, Shuaihang Yuan, Baoru Huang, Anh Nguyen, Mengyu Wang, Anthony Tzes, Yi Fang
机构
*
Embodied AI and Robotics (AIR) Lab, New York University, New York, USA and NYUAD Center for Artificial Intelligence and Robotics, New York University Abu Dhabi, Abu Dhabi, UAE(纽约大学Embodied AI和机器人实验室及纽约大学阿布扎赫尔人工智能与机器人中心)
;
Harvard AI and Robotics Lab, Harvard University, Boston, USA(哈佛大学人工智能与机器人实验室)
;
Department of Computer Science, University College London, London, UK(伦敦大学学院计算机科学系)
;
Department of Computer Science, University of Liverpool, UK(利物浦大学计算机科学系)
;
NYUAD Center for Artificial Intelligence and Robotics, New York University Abu Dhabi, Abu Dhabi, UAE(纽约大学阿布扎赫尔人工智能与机器人中心)
AgenticRAG: Tool-Augmented Foundation Models for Zero-Shot Explainable Recommender Systems
Bo Ma, Hang Li, ZeHua Hu, XiaoFan Gui, LuYao Liu, Simon Liu
机构
*
Department of Software \& Microelectronics Peking University Beijing, China
;
Department of Software \& Microelectronics Peking University Beijing, China hangli\
;
Department of Software \& Microelectronics Peking University Beijing, China zehua\
;
Department of Software \& Microelectronics Peking University Beijing, China xiaofan\
;
Economic Law School China University of Political Science
;
School of Computer Science Peking University Beijing, China
机构
*
McGill University(麦吉尔大学)
;
Massachusetts Institute of Technology(麻省理工学院)
;
Shanghai Jiao Tong University(上海交通大学)
;
The Hong Kong Polytechnic University(香港理工大学)
;
University of Florida(佛罗里达大学)
;
The University of Hong Kong(香港大学)
机构
*
Department of Computer Vision, Mohamed Bin Zayed University of Artificial Intelligence(计算机视觉系,Mohamed Bin Zayed人工智能大学)
;
School of Information Science and Technology, University of Science and Technology of China(信息科学与技术学院,中国科学技术大学)
;
Bytedance Seed(字节跳动种子)
;
School of Computer Science, Qinghai Normal University(计算机科学学院,青海师范大学)
Thinking with Sound: Audio Chain-of-Thought Enables Multimodal Reasoning in Large Audio-Language Models
Zhen Xiong, Yujun Cai, Zhecheng Li, Junsong Yuan, Yiwei Wang
机构
*
University of Southern California(南加州大学)
;
University of Queensland(昆士兰大学)
;
University of California, San Diego(加州大学圣地亚哥分校)
;
University of Buffalo(布法罗大学)
;
University of California, Merced(加州大学默塞德分校)
机构
*
School of Cyber Science and Technology, SUN YAT-SEN UNIVERSITY(中山大学计算机科学与技术学院)
;
School of Computing, National University of Singapore(新加坡国立大学计算机学院)
;
Key Laboratory of Education Informatization for Nationalities, Yunnan Normal University(云南师范大学民族教育信息化重点实验室)
;
SCSE, Beihang University(北航软件学院)