Large Vision-Language Models Get Lost in Attention
大视觉-语言模型在注意力中迷失
Gongli Xi, Ye Tian, Mengyu Yang, Huahui Yi, Liang Lin, Xiaoshuai Hao, Kun Wang, Wendong Wang
机构
*
School of Cyberspace Security, Beijing University of Posts(信息安全学院,北京邮电大学)
;
State Key Laboratory of Networking and Switching Technology, Beijing University of Posts(网络与交换技术国家重点实验室,北京邮电大学)
;
School of Computer Science (National Pilot Software Engineering School), Beijing University of Posts(计算机科学学院(国家试点软件工程学院),北京邮电大学)
;
Nanyang Technological University, Singapore(新加坡南洋理工大学)
;
Institute of Information Engineering, Chinese Academy of Sciences, Beijing, China(信息工程研究所,中国科学院北京研究院)
Adaptive 3D-RoPE: Physics-Aligned Rotary Positional Encoding for Wireless Foundation Models
自适应3D-RoPE:用于无线基础模型的物理对齐旋转位置编码
Chenyu Zhang, Xinchen Lyu, Chenshan Ren, Shuhan Liu, Qimei Cui
机构
*
National Engineering Research Center for Mobile Network Technologies, Beijing University of Posts and Telecommunications(中国移动网络技术国家工程研究中心,北京邮电大学)
;
Department of Broadband Communication, Pengcheng Laboratory(宽带通信部,鹏城实验室)
;
Key Laboratory of Ethnic Language Intelligent Analysis and Security Governance of MOE, Minzu University of China(教育部民族语言智能分析与安全治理重点实验室,中央民族大学)
;
China Telecom Corporation Limited Gansu Branch(中国电信集团甘肃分公司)
Attention-based multiple instance learning for predominant growth pattern prediction in lung adenocarcinoma wsi using foundation models
基于注意力机制的多实例学习用于肺腺癌全切片中主导生长模式预测
Laura Valeria Perez-Herrera, M. J. Garcia-Gonzalez, Karen Lopez-Linares
机构
*
Vicomtech Foundation, Basque Research and Technology Alliance (BRTA)(Vicomtech基金会,巴斯克研究与技术联盟(BRTA))
;
eHealth Group, Bioengineering Area, Biogipuzkoa Health Research Institute(eHealth集团,生物工程领域,Biogipuzkoa健康研究中心)
机构
*
Department of Computer Science and Information Engineering, National Taiwan University, Taiwan(国立台湾大学计算机科学与资讯工程学系)
;
Institute of Information Science, Academia Sinica, Taiwan(台湾“中央研究院”信息科学研究所)
;
AI Research Center (AINTU), National Taiwan University, Taiwan(国立台湾大学人工智能研究中心)
CommentsAccepted to IEEE IGARSS 2025. The final version is available in the Proceedings of the IEEE International Geoscience and Remote Sensing Symposium (IGARSS) 2025
When Numbers Start Talking: Implicit Numerical Coordination Among LLM-Based Agents
当数字开始说话:基于大语言模型的智能体间的隐性数字协调
Alessio Buscemi, Daniele Proverbio, Alessandro Di Stefano, The-Anh Han, German Castignani, Pietro Liò
机构
*
Luxembourg Institute of Science and Technology(卢森堡科学与技术研究所)
;
University of Trento(特伦托大学)
;
Teesside University(泰赛大学)
;
University of Cambridge(剑桥大学)
TTL: Test-time Textual Learning for OOD Detection with Pretrained Vision-Language Models
TTL: 用于基于预训练视觉-语言模型的分布外检测的测试时文本学习
Jinlun Ye, Jiang Liao, Runhe Lai, Xinhua Lu, Jiaxin Zhuang, Zhiyong Gan, Ruixuan Wang
机构
*
Sun Yat-sen University(中山大学)
;
China United Network Communications Corporation Limited Guangdong Branch(中国联合网络通信集团有限公司广东分公司)
;
Peng Cheng Laboratory(鹏城实验室)
;
Hong Kong University of Science and Technology(香港科技大学)
;
Key Laboratory of Machine Intelligence and Advanced Computing, MOE(教育部机器智能与高级计算重点实验室)
Hijacking Large Audio-Language Models via Context-Agnostic and Imperceptible Auditory Prompt Injection
通过上下文无关且不可察觉的音频提示注入劫持大型音频-语言模型
Meng Chen, Kun Wang, Li Lu, Jiaheng Zhang, Tianwei Zhang
机构
*
The State Key Laboratory of Blockchain and Data Security, Zhejiang University(区块链与数据安全国家重点实验室,浙江大学)
;
Hangzhou High-Tech Zone (Binjiang) Institute of Blockchain and Data Security(杭州高新技术区(滨江)区块链与数据安全研究院)
;
Nanyang Technological University(南洋理工大学)
;
National University of Singapore(新加坡国立大学)
机构
*
Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)
;
School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络空间安全学院)
;
National Computer Network Emergency Response Technical Team, Coordination Center of China (CNCERT/CC)(国家计算机网络应急技术处理协调中心)
CommentsAccepted in 34th ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering (FSE Companion 26)
MARL-GPT: Foundation Model for Multi-Agent Reinforcement Learning
MARL-GPT:多智能体强化学习的基础模型
Maria Nesterova, Mikhail Kolosov, Anton Andreychuk, Egor Cherepanov, Oleg Bulichev, Alexey Kovalev, Konstantin Yakovlev, Aleksandr Panov, Alexey Skrynnik
机构
*
MIRAI \& Innopolis University Moscow Russia
;
MIRAI \& Innopolis University
Dependency-Guided Parallel Decoding in Discrete Diffusion Language Models
基于依赖性的并行解码在离散扩散语言模型中
Liran Ringel, Ameen Ali, Yaniv Romano
机构
*
Department of Computer Science, Technion – Israel Institute of Technology(以色列理工学院计算机科学系)
;
Blavatnik School of Computer Science and AI, Tel Aviv, Israel(特拉维夫布拉瓦特尼克计算机科学与人工智能学院)
;
Department of Electrical and Computer Engineering, Technion – Israel Institute of Technology(以色列理工学院电气与计算机工程系)
Phase transition on a context-sensitive random language model with short range interactions
具有短程相互作用的上下文敏感随机语言模型的相变
Yuma Toji, Jun Takahashi, Vwani Roychowdhury, Hideyuki Miyahara
机构
*
Graduate School of Information Science and Technology, Hokkaido University(北海道大学信息科学与技术研究生院)
;
Institute for Solid State Physics, The University of Tokyo(东京大学固体物理研究所)
A chemical language model for reticular materials design
一种用于网状材料设计的化学语言模型
Dhruv Menon, Vivek Singh, Xu Chen, Mohammad Reza Alizadeh Kiapi, Ivan Zyuzin, Hamish W. Macleod, Nakul Rampal, William Shepard, Omar M. Yaghi, David Fairen-Jimenez
机构
*
Department of Chemical Engineering & Biotechnology, University of Cambridge(化学工程与生物技术系,剑桥大学)
;
Department of Chemistry, University of California – Berkeley(化学系,加州大学伯克利分校)
;
Bakar Institute of Digital Materials for the Planet, Berkeley, CA(为地球的数字材料研究所,伯克利,CA)
;
KACST–UC Berkeley Center of Excellence for Nanomaterials for Clean Energy Applications, King Abdulaziz City for Science and Technology(清洁能源应用纳米材料卓越中心,国王阿卜杜勒-阿齐兹城市科学与技术中心)
;
Synchrotron SOLEIL-UR1, L’Orme des Merisiers, Départementale 128, 91190 Saint-Aubin(SOLEIL-UR1同步辐射光源,L’Orme des Merisiers,Départementale 128,91190 Saint-Aubin)