Toward Universal and Transferable Jailbreak Attacks on Vision-Language Models
迈向通用且可迁移的视觉-语言模型劫持攻击
Kaiyuan Cui, Yige Li, Yutao Wu, Xingjun Ma, Sarah Erfani, Christopher Leckie, Hanxun Huang
机构
*
School of Computing and Information Systems, The University of Melbourne, Australia(墨尔本大学计算机与信息系)
;
School of Computing and Information Systems, Singapore Management University, Singapore(新加坡管理大学计算机与信息系)
;
School of Information Technology, Deakin University, Australia(德肯大学信息科技系)
;
Institute of Trustworthy Embodied AI, Fudan University, China(复旦大学可信具身人工智能研究所)
机构
*
School of Computer Science and Technology, Beijing Jiaotong University(北京交通大学计算机科学与技术学院)
;
Shanghai Key Lab of Intell. Info. Processing, School of CS, Fudan University(复旦大学计算机学院智能信息处理重点实验室)
;
Department of Mathematics and Applications, University of Naples Federico II(那不勒斯费德里克二世大学应用数学系)
;
State Key Laboratory of Advanced Rail Autonomous Operation, Beijing Jiaotong University(北京交通大学先进轨道交通自主运行国家重点实验室)
Evaluating Large Vision-language Models for Surgical Tool Detection
评估大型视觉-语言模型用于手术工具检测
Nakul Poudel, Richard Simon, Cristian A. Linte
机构
*
Rochester Institute of Technology(罗切斯特理工大学)
;
Center for Imaging Science(成像科学中心)
;
Center for Imaging Science and Biomedical Engineering(成像科学与生物医学工程中心)
Crafting Adversarial Inputs for Large Vision-Language Models Using Black-Box Optimization
为大型视觉-语言模型设计对抗输入使用黑盒优化
Jiwei Guan, Haibo Jin, Haohan Wang
机构
*
School of Computing, Macquarie University(麦考瑞大学计算机学院)
;
School of Information Sciences, University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校信息科学学院)
A Two-Stage Globally-Diverse Adversarial Attack for Vision-Language Pre-training Models
面向视觉-语言预训练模型的双阶段全局多样化对抗攻击
Wutao Chen, Huaqin Zou, Chen Wan, Lifeng Huang
机构
*
Department of Computer Science and Technology, Shantou University(汕头大学计算机科学与技术系)
;
College of Mathematics and Informatics, South China Agricultural University(华南农业大学数学与信息学院)
Can Vision-Language Models Understand Construction Workers? An Exploratory Study
视觉-语言模型能理解建筑工人吗?一项探索性研究
Hieu Bui, Nathaniel E. Chodosh, Arash Tavakoli
机构
*
Department of Electrical and Computer Engineering(电气与计算机工程系)
;
Villanova University(维拉诺瓦大学)
;
Department of Computing Sciences(计算科学系)
;
Department of Civil and Environmental Engineering(土木与环境工程系)
机构
*
Harbin Institute of Technology, Shenzhen, China(哈尔滨工业大学(深圳))
;
Hong Kong Baptist University, China(香港 Baptist大学)
;
City University of Hong Kong, China(香港城市大学)
MedGround: Bridging the Evidence Gap in Medical Vision-Language Models with Verified Grounding Data
MedGround: 通过验证的 grounding 数据弥合医学视觉-语言模型中的证据差距
Mengmeng Zhang, Xiaoping Wu, Hao Luo, Fan Wang, Yisheng Lv
机构
*
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
DAMO Academy, Alibaba Group(阿里巴巴集团 DAMO 院)
;
Hupan Lab, Zhejiang Province(浙江省 Hupan 实验室)