Comments18 pages, 3 figures. To be published in proceedings of the 25th International Conference on Autonomous Agents and Multi-Agent Systems (AAMAS 2026). This is a full version that includes the supplementary material
Omni-Safety under Cross-Modality Conflict: Vulnerabilities, Dynamics Mechanisms and Efficient Alignment
跨模态冲突下的全方位安全:漏洞、动态机制和高效对齐
Kun Wang, Zherui Li, Zhenhong Zhou, Yitong Zhang, Yan Mi, Kun Yang, Yiming Zhang, Junhao Dong, Zhongxiang Sun, Qiankun Li, Yang Liu
机构
*
Nanyang Technological University(南洋理工大学)
;
Beijing University of Posts and Telecommunications(北京邮电大学)
;
Tsinghua University(清华大学)
;
Fudan University(复旦大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Renmin University of China(中国人民大学)
Red-teaming the Multimodal Reasoning: Jailbreaking Vision-Language Models via Cross-modal Entanglement Attacks
多模态推理的红队测试:通过跨模态纠缠攻击劫持视觉-语言模型
Yu Yan, Sheng Sun, Shengjia Cheng, Teli Liu, Mingfeng Li, Min Liu
机构
*
Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
People’s Public Security University of China(中国人民公安大学)
from Benign import Toxic: Jailbreaking the Language Model via Adversarial Metaphors
从无害到有害:通过对抗隐喻 jailbreak 语言模型
Yu Yan, Sheng Sun, Zenghao Duan, Teli Liu, Min Liu, Zhiyi Yin, Jingyu Lei, Qi Li
机构
*
Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
People’s Public Security University of China(中国人民公安大学)
;
Tsinghua University(清华大学)
Copyright Detective: A Forensic System to Evidence LLMs Flickering Copyright Leakage Risks
版权侦探:一种用于证据LLM闪烁版权泄露风险的取证系统
Guangwei Zhang, Jianing Zhu, Cheng Qian, Neil Gong, Rada Mihalcea, Zhaozhuo Xu, Jingrui He, Jiaqi Ma, Yun Huang, Chaowei Xiao, Bo Li, Ahmed Abbasi, Dongwon Lee, Heng Ji, Denghui Zhang
机构
*
Pine AI
;
The University of Texas at Austin(德克萨斯大学奥斯汀分校)
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
Duke University(杜克大学)
;
University of Michigan(密歇根大学)
;
Stevens Institute of Technology(史蒂文斯理工学院)
;
Johns Hopkins University(约翰霍普金斯大学)
;
University of Notre Dame(圣母大学)
;
The Pennsylvania State University(宾夕法尼亚州立大学)
A Multimodal Manufacturing Safety Chatbot: Knowledge Base Design, Benchmark Development, and Evaluation of Multiple RAG Approaches
多模态制造安全聊天机器人:知识库设计、基准开发及多种RAG方法的评估
Ryan Singh, Austin Hamilton, Amanda White, Michael Wise, Ibrahim Yousif, Arthur Carvalho, Zhe Shan, Reza Abrisham Baf, Mohammad Mayyas, Lora A. Cavuoto, Fadel M. Megahed
机构
*
Farmer School of Business, Miami University(Miami大学农业商学院)
;
Department of Computer Science and Software Engineering, Miami University(Miami大学计算机科学与软件工程系)
;
Department of Engineering Technology, Miami University(Miami大学工程技术系)
;
Department of Mechanical and Manufacturing Engineering, Miami University(Miami大学机械与制造工程系)
;
Department of Industrial and Systems Engineering, University at Buffalo(University at Buffalo工业与系统工程系)
Safety with Agency: Human-Centered Safety Filter with Application to AI-Assisted Motorsports
安全与代理:以人为中心的安全过滤器及其在AI辅助赛车中的应用
Donggeon David Oh, Justin Lidard, Haimin Hu, Himani Sinhmar, Elle Lazarski, Deepak Gopinath, Emily S. Sumner, Jonathan A. DeCastro, Guy Rosman, Naomi Ehrich Leonard, Jaime Fernández Fisac
机构
*
Department of Electrical and Computer Engineering, Princeton University(普林斯顿大学电气与计算机工程系)
;
Department of Mechanical and Aerospace Engineering, Princeton University(普林斯顿大学机械与航空航天工程系)
;
Toyota Research Institute(丰田研究院)
CommentsTo be published in IEEE International Workshop on Decentralized Physical Infrastructure Networks 2025, in conjunction with ICBC'25. 7 pages. 3 figures