Toward Universal and Transferable Jailbreak Attacks on Vision-Language Models
迈向通用且可迁移的视觉-语言模型劫持攻击
Kaiyuan Cui, Yige Li, Yutao Wu, Xingjun Ma, Sarah Erfani, Christopher Leckie, Hanxun Huang
机构
*
School of Computing and Information Systems, The University of Melbourne, Australia(墨尔本大学计算机与信息系)
;
School of Computing and Information Systems, Singapore Management University, Singapore(新加坡管理大学计算机与信息系)
;
School of Information Technology, Deakin University, Australia(德肯大学信息科技系)
;
Institute of Trustworthy Embodied AI, Fudan University, China(复旦大学可信具身人工智能研究所)
Comments10 pages, To Be Published In Proceedings Of The 1st IEEE Workshop on Healthcare and Medical Device Security, Privacy, Resilience, and Trust (IEEE HMD-SPiRiT), Accepted & Presented At The 7th IEEE International Conference on Trust, Privacy & Security in Intelligent Systems, and Applications (IEEE TPS 2025) on Nov. 11, 2025 in Pittsburgh, PA, USA