Security Risk of Misalignment between Text and Image in Multi-modal Model
机构 * Xidian University(西安电子科技大学) ; Shanghai Jiaotong University(上海交通大学)
专题命中 多模态生成 :multi-modal(title,abstract);分类 cs.CV、cs.AI
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
机构 * Xidian University(西安电子科技大学) ; Shanghai Jiaotong University(上海交通大学)
专题命中 多模态生成 :multi-modal(title,abstract);分类 cs.CV、cs.AI
机构 * BAAI(百度人工智能研究院)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV
Comments project page: https://emu.world
机构 * National University of Singapore(新加坡国立大学) ; University of Maryland, College Park(马里兰大学 College Park 分校) ; University of California, Los Angeles(加州大学洛杉矶分校)
专题命中 多模态生成 :any-to-any(title);multimodal(abstract)
Comments 44 pages, 9 figures, 13 tables, paper accepted by NeurIPS 2025
机构 * Electronics and Telecommunications Research Institute, Republic of Korea(韩国电子电信研究院) ; POSTECH ; Sungkyunkwan University(全南大学)
专题命中 多模态生成 :multimodal(abstract);cross-modal(abstract);分类 cs.CL、cs.AI
Comments NeurIPS 2025, 38 pages, 8 figures
机构 * Shanghai Jiao Tong University(上海交通大学) ; National University of Singapore(国立新加坡大学) ; Shanghai AI Lab(上海人工智能实验室)
专题命中 多模态生成 :multi-modal(title)
专题命中 多模态生成 :multimodal(abstract);cross-modal(abstract);分类 cs.CV
Comments Accepted to IEEE Transactions on Multimedia (TMM)
专题命中 多模态生成 :multi-modal(abstract)
Comments Accepted by the 48th IEEE/ACM International Conference on Software Engineering (ICSE 2026)