MARE: Multimodal Alignment and Reinforcement for Explainable Deepfake Detection via Vision-Language Models
MARE: 多模态对齐与强化学习用于通过视觉-语言模型的可解释深度伪造检测
Wenbo Xu, Wei Lu, Xiangyang Luo, Jiantao Zhou
机构
*
School of Computer Science and Engineering, MoE Key Laboratory of Information Technology, Guangdong Province Key Laboratory of Information Security Technology, Sun Yat-sen University, Guangzhou 510006, China(计算机科学与工程学院,信息技术MOE实验室,广东省信息安全技术重点实验室,中山大学,广州510006,中国)
;
State Key Laboratory of Mathematical Engineering and Advanced Computing(数学工程与先进计算国家重点实验室)
;
Department of Computer and Information Science, University of Macau.(计算机与信息科学系,澳门大学)
机构
*
College of Energy Engineering, Zhejiang University(浙江大学能源工程学院)
;
Polytechnic Institute, Zhejiang University(浙江大学 polytechnic 院)
;
Shanghai Institute for Advanced Study, Zhejiang University(浙江大学上海研究院)
Metis-SPECS: Decoupling Multimodal Learning via Self-distilled Preference-based Cold Start
Metis-SPECS: 通过基于偏好自我蒸馏的冷启动解耦多模态学习
Kun Chen, Peng Shi, Haibo Qiu, Zhixiong Zeng, Siqi Yang, Wenji Mao, Lin Ma
机构
*
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
MAIS, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所MAIS)
;
Meituan(美团)
Clipping-Free Policy Optimization for Large Language Models
无需裁剪的策略优化用于大语言模型
Ömer Veysel Çağatan, Barış Akgün, Gözde Gül Şahin, Xuandong Zhao
机构
*
KUIS AI Center, Koç University, Istanbul, Türkiye(Koç大学)
;
Koç University, Istanbul, Türkiye(Koç大学)
;
University of California, Berkeley, CA, USA(加州大学伯克利分校)
机构
*
Beijing Key Laboratory of Safe AI and Superalignment(北京安全人工智能与超对齐重点实验室)
;
Beijing Institute of AI Safety and Governance(北京人工智能安全与治理研究院)
;
Brain-inspired Cognitive AI Lab, Institute of Automation, Chinese Academy of Sciences(脑启发认知人工智能实验室,中国科学院自动化研究所)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Long-term AI(长期人工智能)
A Systematic Literature Review on LLM Defenses Against Prompt Injection and Jailbreaking: Expanding NIST Taxonomy
针对提示注入和劫持攻击的LLM防御措施系统文献综述:扩展NIST分类
Pedro H. Barcha Correia, Ryan W. Achjian, Diego E. G. Caetano de Oliveira, Ygor Acacio Maria, Victor Takashi Hayashi, Marcos Lopes, Charles Christian Miers, Marcos A. Simplicio
In Vino Veritas and Vulnerabilities: Examining LLM Safety via Drunk Language Inducement
葡萄酒中的真理与漏洞:通过醉酒语言诱导检验LLM安全性
Anudeex Shetty, Aditya Joshi, Salil S. Kanhere
机构
*
School of Computer Science and Engineering, UNSW Sydney(计算机科学与工程学院,新南威尔士大学悉尼分校)
;
School of Computing and Information System, the University of Melbourne(计算与信息系统学院,墨尔本大学)
机构
*
Tencent Youtu Lab
;
East China University of Science
;
Shenzhen University
;
Hong Kong University of Science
;
Department of XXX, University of YYY, Location, Country
;
School of ZZZ, Institute of WWW, Location, Country
机构
*
Gaoling School of Artificial Intelligence\ University of China Beijing China
;
School of Information Technology
;
Search Applications Department, Tencent Beijing China
;
Beijing University of Posts
;
Gaoling School of Artificial Intelligence\ University of China
;
Search Applications Department, Tencent
机构
*
Renmin University of China(中国人民大学)
;
Gaoling School of Artificial Intelligence(北京人工智能学院)
;
Amap, Alibaba Group(阿里集团阿里的地图部门)
;
School of Computer Science & Technology(计算机科学与技术学院)