Local Margin Restoration for Test-Time Adaptation of Vision-Language Models
用于视觉-语言模型测试时适应的局部间隔恢复
Yan Huang, Guowei Wang, Xu Wang, Kangjun Liu, Xin Lin
机构
*
Guangzhou University(广州大学)
;
The Second Affiliated Hospital of Guangzhou University of Chinese Medicine(广州中医药大学第二附属医院)
;
Jinan University(暨南大学)
;
Pengcheng Laboratory(鹏城实验室)
Unifying Adversarially Robust Model Experts in Vision-Language Models
统一视觉-语言模型中的对抗鲁棒模型专家
Nguyen Duc Thai, Junhao Dong, Sua Qi Rong, Hua Yu, Yew-Soon Ong
机构
*
College of Computing and Data Science, Nanyang Technological University(南洋理工大学计算与数据科学学院)
;
Center for Frontier AI Research, Agency for Science, Technology and Research (A*STAR)(新加坡科技研究局前沿人工智能研究中心)
机构
*
School of Cyber Science and Engineering, Huazhong University of Science and Technology(华中科技大学网络空间安全学院)
;
College of Computer Science, Chongqing University(重庆大学计算机科学学院)
;
School of Software and engineering, Huazhong University of Science and Technology(华中科技大学软件工程学院)
;
School of Information and Communication Technology, Griffith University(格里菲斯大学信息与通信技术学院)
专题命中
幻觉与鲁棒性
:vision-language model(title);vision language model(abstract);分类 cs.CV
Towards Fast and Effective Long Video Understanding of Multimodal Large Language Models via Adaptive Quasi-Gaussian Sampling
面向多模态大语言模型的长视频快速有效理解:自适应准高斯采样
Kun Zhang, Chenxin Fang, Tao Chen, Baiyang Song, Yunhang Shen, Yiyi Zhou, Rongrong Ji
机构
*
Key Laboratory of Multimedia Trusted Perception and Efficient Computing, Ministry of Education of China, Xiamen University(厦门大学多媒体可信感知与高效计算教育部重点实验室)
专题命中
幻觉与鲁棒性
:multimodal large language model(title,abstract);分类 cs.CV
Adversarial Diffusion Across Modalities: A Fusion Survey of Attacks, Defenses, and Evaluation for Text, Vision, and Vision-Language Models
跨模态对抗扩散:文本、视觉与视觉-语言模型的攻击、防御与评估融合综述
Abrar Alotaibi, Moataz Ahmed
机构
*
Information and Computer Science Department, King Fahd University of Petroleum & Minerals(国王法赫德石油矿物大学信息与计算机科学系)
;
College of Computer Science and Information Technology, Imam Abdulrahman Bin Faisal University(伊玛目阿卜杜勒拉赫曼·本·法伊塞尔大学计算机科学与信息科技学院)
;
SDAIA-KFUPM Joint Research Center for Artificial Intelligence, King Fahd University of Petroleum & Minerals(SDAIA-KFUPM人工智能联合研究中心,国王法赫德石油矿物大学)
MLLMs Get It Right, Then Get It Wrong: Tracing and Correcting Late-Layer Textual Bias
MLLMs 先正确后错误:追踪并纠正后层文本偏见
Xingming Li, Ao Cheng, Qiyao Sun, Xixiang He, Xuanyu Ji, Runke Huang, Qingyong Hu
机构
*
National University of Defense Technology(国防科技大学)
;
Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
Intelligent Game and Decision Lab(智能博弈与决策实验室)
专题命中
幻觉与鲁棒性
:grounding(abstract);multimodal large language model(abstract);MLLM(abstract_cn);分类 cs.CV
On the Adversarial Robustness of Multimodal LLM Judges
多模态大语言模型评判器的对抗鲁棒性
Zihan Wang, Guansong Pang, Zelin Liu, Wenjun Miao, Jin Zheng, Xiao Bai
机构
*
School of Computer Science and Engineering, Beihang University(北京航空航天大学计算机科学与工程学院)
;
State Key Laboratory of Virtual Reality Technology and System, Beihang University(北京航空航天大学虚拟现实技术与系统国家重点实验室)
;
State Key Laboratory of Software Development Environment, Jiangxi Research Institute, Beihang University(北京航空航天大学江西研究院软件开发环境国家重点实验室)
;
School of Computing and Information Systems, Singapore Management University(新加坡管理大学计算机与信息系统学院)
专题命中
幻觉与鲁棒性
:MLLM(abstract,abstract_cn);multimodal large language model(abstract);分类 cs.CV
机构
*
Advanced Technologies Application Center (CENATAV)(先进技术应用中心(CENATAV))
;
Centro de Sistemas Complejos, Facultad de Física, Universidad de La Habana(哈瓦那大学物理学院复杂系统中心)
专题命中
幻觉与鲁棒性
:MLLM(abstract,abstract_cn);multimodal large language model(abstract);分类 cs.CV