The Safety Reminder: A Soft Prompt to Reactivate Delayed Safety Awareness in Vision-Language Models
Peiyuan Tang, Haojie Xin, Xiaodong Zhang, Jun Sun, Qin Xia, Zijiang Yang
机构
*
School of Computer Science and Technology, Xi’an Jiaotong University(西安交通大学计算机科学与技术学院)
;
School of Computer Science and Technology, University of Science and Technology of China(中国科学技术大学计算机科学与技术学院)
;
School of Computing and Information Systems, Singapore Management University(新加坡管理学院计算与信息系统学院)
Failures to Find Transferable Image Jailbreaks Between Vision-Language Models
Rylan Schaeffer, Dan Valentine, Luke Bailey, James Chua, Cristóbal Eyzaguirre, Zane Durante, Joe Benton, Brando Miranda, Henry Sleight, John Hughes, Rajashree Agrawal, Mrinank Sharma, Scott Emmons, Sanmi Koyejo, Ethan Perez
The Illusion of Progress? A Critical Look at Test-Time Adaptation for Vision-Language Models
Lijun Sheng, Jian Liang, Ran He, Zilei Wang, Tieniu Tan
机构
*
University of Science and Technology of China(中国科学技术大学)
;
NLPR & MAIS, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
RadAgent: A tool-using AI agent for stepwise interpretation of chest computed tomography
RadAgent:一种用于胸部CT逐步解读的工具型AI智能体
Mélanie Roschewitz, Kenneth Styppa, Yitian Tao, Jiwoong Sohn, Jean-Benoit Delbrouck, Benjamin Gundersen, Nicolas Deperrois, Christian Bluethgen, Julia E. Vogt, Bjoern Menze, Farhad Nooralahzadeh, Michael Krauthammer, Michael Moor
机构
*
Department of Biosystems Science and Engineering, ETH Zurich(生物系统科学与工程系,苏黎世联邦理工学院)
;
ETH AI Center, Zurich(ETH人工智能中心,苏黎世)
;
Department of Computer Science, ETH Zurich(计算机科学系,苏黎世联邦理工学院)
;
Faculty of Computer Science and Mathematics, Heidelberg University(计算机科学与数学学院,海德堡大学)
;
Stanford Center for Artificial Intelligence in Medicine and Imaging, Stanford University(斯坦福大学人工智能在医学和影像中的中心)
;
Department of Radiology, Stanford University(放射科,斯坦福大学)
;
Department of Quantitative Biomedicine, University of Zurich(定量生物医学系,苏黎世大学)
;
Institute of Computer Science, Zurich University of Applied Sciences(应用科学大学计算机科学研究所)
机构
*
Research Intern, Department of Mechanical and Aerospace Engineering, George Washington University(乔治华盛顿大学机械与航空航天工程系研究实习生)
;
Ph.D. Student, Department of Mechanical and Aerospace Engineering, George Washington University(乔治华盛顿大学机械与航空航天工程系博士生)
;
Undergraduate Student, Aerospace Program, University of California, Berkeley(加州大学伯克利分校航空航天项目本科生)
;
Full Professor, Department of Electrical Engineering and Computer Science, University of California, Berkeley(加州大学伯克利分校电气工程与计算机科学系教授)
;
Ph.D. Student, Department of Computer Science, George Washington University(乔治华盛顿大学计算机科学系博士生)
;
Associate Professor, Department of Mechanical and Aerospace Engineering, George Washington University(乔治华盛顿大学机械与航空航天工程系副教授)
Unveiling the Fragility of Vision-Language Models: Multi-Modal Adversarial Synergy via Texture-Constrained Perturbations and Cross-Modal Optimization
揭示视觉-语言模型的脆弱性:通过纹理约束扰动和跨模态优化的多模态对抗协同
Xiang Fang, Wanlong Fang, Changshuo Wang
机构
*
School of Software Engineering, Huazhong University of Science and Technology(华中科技大学软件学院)
;
Nanyang Technological University, Singapore(新加坡南洋理工大学)
;
University College London(伦敦大学学院)
机构
*
Shanghai Key Laboratory of Multidimensional Information Processing, East China Normal University(多维信息处理上海市重点实验室,东华大学)
;
Zhongguancun Academy(中关村学院)
专题命中
幻觉与鲁棒性
:vision-language model(title);vision language model(abstract);grounding(abstract);分类 cs.CV、cs.AI