CommentsThe authors request withdrawal of this article. This version was submitted in error. Compared to the intended final version, it contains inaccuracies and fails to accurately reflect the authors' work and conclusions
机构
*
College of Computer Science, Sichuan University(四川大学计算机科学学院)
;
Southwest China Institute of Electronic Technology(西南中国电子技术研究所)
;
National Key Laboratory of Fundamental Algorithms(国家基础算法重点实验室)
;
Models for Engineering Simulation, Sichuan University(工程模拟模型研究所,四川大学)
Text Before Vision: Staged Knowledge Injection Matters for Agentic RLVR in Ultra-High-Resolution Remote Sensing Understanding
文本优先于视觉:针对超高清遥感理解的代理强化学习在超高清遥感理解中的知识注入至关重要
Fengxiang Wang, Mingshuo Chen, Yueying Li, Yajie Yang, Yuhao Zhou, Di Wang, Yifan Zhang, Haoyu Wang, Haiyan Zhao, Hongda Sun, Long Lan, Jun Song, Yulin Wang, Jing Zhang, Wenlong Zhang, Bo Du
机构
*
National University of Defense Technology, China(国防科技大学)
;
Beijing University of Posts and Telecommunications, China(北京邮电大学)
;
University of the Chinese Academy of Sciences, China(中国科学院大学)
;
Sichuan University, China(四川大学)
;
Wuhan University, China(武汉大学)
;
Chinese Academy of Science, China(中国科学院)
;
Tsinghua University, China(清华大学)
;
Shanghai Artificial Intelligence Laboratory, China(上海人工智能实验室)
;
Renmin University of China, China(中国人民大学)
机构
*
Key Laboratory of Child Development and Learning Science (Ministry of Education), School of Biological Sciences and Medical Engineering, Southeast University(儿童发展与学习科学重点实验室(教育部),生物科学与医学工程学院,东南大学)
;
School of Computer Science and Engineering, Nanjing University of Science and Technology(计算机科学与工程学院,南京理工大学)
;
School of Computer Science and Engineering, Nanyang Technological University(计算机科学与工程学院,南洋理工大学)
Pyramid Token Pruning for High-Resolution Large Vision-Language Models via Region, Token, and Instruction-Guided Importance
通过区域、令牌和指令引导的重要性进行高分辨率大视觉-语言模型的金字塔令牌修剪
Yuxuan Liang, Xu Li, Xiaolei Chen, Yi Zheng, Haotian Chen, Bin Li, Xiangyang Xue
机构
*
Shanghai Key Laboratory of Intelligent Information Processing(上海智能信息处理重点实验室)
;
College of Computer Science and Artificial Intelligence(计算机科学与人工智能学院)
Enhancing spatial hearing with cochlear implants: exploring the role of AI, multimodal interaction and perceptual training
通过 cochlear implants 增强空间听觉:探索人工智能、多模态交互和感知训练的作用
Lorenzo Picinali, Robert Baumgartner, Valerie Gaveau, Antonino Greco, Stefanie Liebe, Paul Oomen, Christoph Braun
机构
*
Acoustics Research Institute, Austrian Academy of Sciences(奥地利科学院声学研究所)
;
Centre de Recherche en Neurosciences de Lyon Inserm(里昂神经科学研究中心(Inserm))
;
Department of Neural Dynamics and Magnetoencephalography, Hertie Institute for Clinical Brain Research, University of Tübingen(图宾根大学神经动力学与脑磁图部门)
;
Centre for Integrative Neuroscience, University of Tübingen(图宾根大学整合神经科学中心)
;
NEMO Labs Nonprofit Kft., Budapest(布达佩斯NEMO实验室(非营利公司))
Breaking Data Efficiency Dilemma: A Federated and Augmented Learning Framework For Alzheimer's Disease Detection via Speech
突破数据效率困境:一种联邦学习与增强学习框架用于通过语音检测阿尔茨海默病
Xiao Wei, Bin Wen, Yuqin Lin, Kai Li, Mingyang gu, Xiaobao Wang, Longbiao Wang, Jianwu Dang
机构
*
Tianjin Key Laboratory of Cognitive Computing and Application(认知计算与应用天津重点实验室)
;
College of Intelligence and Computing, Tianjin University(智能计算学院,天津大学)
;
Shenzhen Institutes of Advanced Technology, Chinese Academy of Sciences(深圳先进技术研究院,中国科学院)
;
College of Computer and Data Science, Fuzhou University(计算机与数据科学学院,福州大学)
;
Huiyan Technology (Tianjin) Co., Ltd(慧研科技(天津)有限公司)
How Do Lexical Senses Correspond Between Spoken German and German Sign Language?
德语口语中的词义如何与德国手语对应?
Melis Çelikkol, Wei Zhao
机构
*
Institute for Computational Linguistics, University of Heidelberg(计算语言学研究所,海德堡大学)
;
Department of Computing Science, University of Aberdeen(计算机科学系,阿伯丁大学)
C^2ROPE: Causal Continuous Rotary Positional Encoding for 3D Large Multimodal-Models Reasoning
C^2ROPE: 3D 大多模态模型推理中的因果连续旋转位置编码
Guanting Ye, Qiyan Zhao, Wenhao Yu, Xiaofeng Zhang, Jianmin Ji, Yanyong Zhang, Ka-Veng Yuen
机构
*
State Key Laboratory of Internet of Things for Smart City, University of Macau(物联网智能城市国家重点实验室,澳门大学)
;
Department of Automation, Shanghai Jiaotong University(上海交通大学自动化系)
;
Institute of Advanced Technology, University of Science and Technology of China(中国科学技术大学先进技术研究院)
;
School of Computer Science and Technology, USTC(中国科学技术大学计算机科学与技术学院)
;
School of Artificial Intelligence and Data Science, USTC(中国科学技术大学人工智能与数据科学学院)
Simulating the Real World: A Unified Survey of Multimodal Generative Models
模拟现实世界:多模态生成模型的统一综述
Yuqi Hu, Longguang Wang, Xian Liu, Ling-Hao Chen, Yuwei Guo, Yukai Shi, Ce Liu, Anyi Rao, Zeyu Wang, Hui Xiong
机构
*
Thrust of Artificial Intelligence, The Hong Kong University of Science and Technology (Guangzhou)(人工智能前沿技术研究所,香港科学与技术大学(广州))
;
Department of Computer Science and Engineering, The Hong Kong University of Science and Technology Hong Kong SAR(计算机科学与工程系,香港科学与技术大学香港特别行政区)
;
MMLab, The Hong Kong University of Science and Technology(多模态实验室,香港科学与技术大学)
;
School of Electronics and Communication Engineering, Shenzhen Campus of Sun Yat-sen University(电子与通信工程学院,中山大学深圳校区)
;
The Chinese University of Hong Kong, Hong Kong, China(香港中文大学,香港,中国)
;
Tsinghua University, Guangdong, China(清华大学,广东,中国)
;
Bosch (China) Investment Co., Ltd., Shanghai, China(博世(中国)投资有限公司,上海,中国)