机构
*
Beijing University of Posts and Telecommunications(北京邮电大学)
;
Tsinghua University(清华大学)
;
Hunan Normal University(湖南师范大学)
;
City University of Hong Kong(香港城市大学)
机构
*
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
National University of Singapore(新加坡国立大学)
;
Southeast University, Nanjing, China(南京东南大学)
;
Wuhan AI Research, Wuhan, China(武汉人工智能研究院)
专题命中
VLM训练与架构
:multimodal large language model(abstract);分类 cs.AI、cs.LG
CommentsPublished in the Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (ACL 2026), Volume 1: Long Papers. 14 pages. Code is available at https://github.com/yueluoshuangtian/PASs-MoE
Journal refProceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), pages 31959--31972, San Diego, California, United States, July 2026. Association for Computational Linguistics
Comments6 pages, 5 figures, 2 tables submitted to 2026 Joint 14th International Conference on Soft Computing and Intelligent Systems and 27th International Symposium on Advanced Intelligent Systems (SCIS&ISIS 2026)
PFAdapter: Hierarchical LoRA Decomposition for Personalized Federated MLLMs
PFAdapter:用于个性化联邦多模态大语言模型的分层LoRA分解
Jing Liu, Kun Yang, Yan Wang, Dingkang Yang, Xiaoshuai Hao, Wei Zhang, Yang Liu, Wei Zhou
机构
*
College of Future Information Technology, Fudan University(复旦大学未来信息技术学院)
;
Division of Natural and Applied Sciences, Duke Kunshan University(昆山杜克大学自然科学与应用科学部)
;
Department of Electrical and Computer Engineering, The University of British Columbia(英属哥伦比亚大学电气与计算机工程系)
;
Ant Group(蚂蚁集团)
;
College of Information Science and Electronic Engineering, Zhejiang University(浙江大学信息科学与电子工程学院)
;
School of Data Science and Engineering, East China Normal University(华东师范大学数据科学与工程学院)
;
College of Intelligent Robotics and Advanced Manufacturing, Fudan University & Fysics AI(复旦大学智能机器人与先进制造学院及复肆智能科技(上海)有限公司)
;
Xiaomi EV, Xiaomi Campus(小米汽车、小米园区)
;
Information and Communications Technology Cluster, Singapore Institute of Technology (SIT)(新加坡理工学院信息通信技术集群)
;
College of Electronic and Information Engineering, Tongji University(同济大学电子与信息工程学院)
;
School of Computer Science and Informatics, Cardiff University(卡迪夫大学计算机科学与信息学院)
专题命中
VLM训练与架构
:multimodal large language model(abstract);分类 cs.AI、cs.LG
MedPMC: A Systematic Framework for Scaling High-Fidelity Medical Multimodal Data for Foundation Models
MedPMC:一种用于为基础模型扩展高保真医学多模态数据的系统框架
Hyunjae Kim, Dain Kim, Pan Xiao, Serina S. Applebaum, Younjoon Chung, Xuguang Ai, Yu Yin, Roy Jiang, Yuexi Du, Yawen Wei, Yiming Kong, Tuo Guo, Zhiyuan Cao, Mengmeng Du, Yuelei Fu, Yan Hu, Rui Shi, Gui Yang, Kevin W. Jin, Yuntian Liu, Yuxuan Tian, Jonathan Marquez, Zhen Chen, Sheng Zhang, Hoifung Poon, Hua Xu, Jaewoo Kang, Qingyu Chen
机构
*
Yale University(耶鲁大学)
;
Korea University(韩国大学)
;
Washington University in St. Louis(圣路易斯华盛顿大学)
;
The University of Queensland(昆士兰大学)
;
The University of Texas Health Science Center at Houston(德克萨斯大学休斯顿健康科学中心)
;
University of Washington(华盛顿大学)
;
Microsoft Research(微软研究院)
专题命中
VLM训练与架构
:multimodal large language model(abstract);分类 cs.CV、cs.LG
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models
InverseScope:用于解释大语言模型的可扩展激活反演
Yifan Luo, Zhennan Zhou, Bin Dong
机构
*
School of Mathematical Science, Peking University(北京大学数学科学学院)
;
School of Science, Westlake University(西湖大学理学院)
;
Beijing International Center for Mathematical Research(北京国际数学研究中心)
;
New Cornerstone Science Laboratory, Peking University(北京大学新基石科学实验室)
ReMAP-PET: Beyond Visual Understanding -- Learning Region-Guided Metabolic Alignment Semantics from Brain PET
ReMAP-PET:超越视觉理解——从脑PET学习区域引导的代谢对齐语义
Dasen Dai, Yanteng Zhang, Shuoqi Li, Yuxiang Wei, Hongjie Yu, Qingxin Zhang, Qizhen Lan, Jagath C. Rajapakse, Vince D. Calhoun
机构
*
The Chinese University of Hong Kong, HKSAR(香港中文大学)
;
TReNDS Center (Georgia State, Georgia Tech, Emory)(TReNDS中心(佐治亚州立大学、佐治亚理工学院、埃默里大学))
;
ShanghaiTech University, Shanghai, P.R.China(上海科技大学)
;
University of California, Berkeley, USA(加州大学伯克利分校)
;
University of Texas Health Science Center at Houston, USA(德克萨斯大学健康科学中心(休斯顿))
;
Nanyang Technological University, Singapore(南洋理工大学)
机构
*
Hong Kong University of Science and Technology (Guangzhou)(香港理工大学(广州))
;
Tsinghua University(清华大学)
;
Nanyang Technological University(南洋理工大学)
;
Renmin University of China(中国人民大学)
;
University of Science and Technology of China(中国科学技术大学)
;
University of Chinese Academy of Sciences(中国科学院大学)
HAT-4D: Lifting Monocular Video for 4D Multi-Object Interactions via Human-Agent Collaboration
HAT-4D: 通过人机协作从单目视频提升4D多物体交互
Jiaxin Li, Yuxiang Wu, Zhenkai Zhang, Xinrui Shi, Haoyuan Wang, Yichen Zhao, Su Linxiang, Chenyang Yu, Mingyu Zhang, Yifan Ding, Boran Wen, Li Zhang, Ruiyang Liu, Yong-Lu Li
机构
*
Shanghai Jiao Tong University(上海交通大学)
;
Shanghai Innovation Institute(上海创新研究院)
;
University of Science and Technology of China(中国科学技术大学)
;
Math Magic
Evaluating the Interpretability of Sparse Autoencoders with Concept Annotations
评估稀疏自编码器与概念标注的可解释性
Jonas Klotz, Cassio F. Dantas, Pallavi Jain, Diego Marcos, Begüm Demir
机构
*
The Berlin Institute for the Foundations of Learning and Data (BIFOLD)(柏林学习与数据基础研究所)
;
Technische Universität Berlin(柏林工业大学)
;
INRAE(法国国家农业、食品与环境研究院)
;
Inria, EVERGREEN(法国国家信息与自动化研究所,EVERGREEN)
;
UMR TETIS, Univ Montpellier(UMR TETIS,蒙彼利埃大学)
专题命中
VLM训练与架构
:vision language model(abstract);分类 cs.CV、cs.AI