TRACE: Textual Relevance Augmentation and Contextual Encoding for Multimodal Hate Detection
专题命中 图文多模态 :multimodal(title,abstract);分类 cs.CV、cs.CL
Comments Accepted to Special Track on AI for Social Impact (AISI) at AAAI 2026
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
专题命中 图文多模态 :multimodal(title,abstract);分类 cs.CV、cs.CL
Comments Accepted to Special Track on AI for Social Impact (AISI) at AAAI 2026
专题命中 图文多模态 :multimodal(title,abstract);分类 cs.CL、cs.AI
专题命中 图文多模态 :multimodal(title,abstract)
Comments Accepted for presentation at CORL 2025. Code, models, and data are available at https://search-tta.github.io/
机构 * PatSnap Co., LTD.(PatSnap公司)
专题命中 图文多模态 :multimodal(abstract);image-text(abstract);分类 cs.CV
机构 * German Research Centre for Artificial Intelligence (DFKI)(德国人工智能研究中心) ; Max Planck Research School for Intelligent Systems (IMPRS-IS)(马克斯·普朗克智能系统研究学校) ; University of Stuttgart(斯图加特大学) ; University Medical Center Gottingen(哥廷根大学医学中心) ; Max Planck Institute for Multidisciplinary Sciences(马克斯·普朗克多学科科学研究所) ; ARC Centre of Excellence for the Mathematical Analysis of Cellular Systems(细胞系统数学分析卓越中心) ; School of Mathematical Sciences, Queensland University of Technology(昆士兰科技大学数学科学学院) ; University of Oldenburg(奥尔登堡大学) ; University of Texas at Austin(德克萨斯大学奥斯汀分校) ; University of California San Diego(加州大学圣地亚哥分校) ; MBZUAI(马克斯·普朗克人工智能研究所) ; ETH Zurich(苏黎世联邦理工学院) ; Stanford University(斯坦福大学)
专题命中 图文多模态 :multi-modal(abstract);cross-modal(abstract)
Comments Accepted at NeurIPS 2025
机构 * Dublin City University(都柏林城市大学)
专题命中 图文多模态 :multimodal(abstract);分类 cs.AI、cs.MM
Comments Accepted for publication at MMM'26. Analysis code can be found here: https://github.com/allie-tran/clip-brittleness
机构 * Karlsruhe Institute of Technology(卡尔斯鲁厄理工学院) ; Carnegie Mellon University(卡内基梅隆大学)
专题命中 音频语音多模态 :audio-visual(title,abstract);分类 cs.CV
专题命中 音频语音多模态 :multimodal(title,abstract)
机构 * Center for Thinking, Language and Communication(思考、语言与交流中心) ; Plaksha University(普拉克斯大学)
专题命中 音频语音多模态 :multi-modal(abstract);分类 cs.CL、cs.AI
机构 * Department of Statistics and Actuarial Science, School of Computing and Data Science, The University of Hong Kong(统计与精算学系,计算与数据科学学院,香港大学)
专题命中 视频多模态 :multimodal(title,abstract);cross-modal(title,abstract);分类 cs.AI
Comments ACL 2025 Findings
机构 * University of Tennessee(田纳西大学)
专题命中 视频多模态 :multi-modal(title)
专题命中 视频多模态 :multimodal(abstract);分类 cs.AI
Comments 185 pages
Journal ref Foundations and Trends in Signal Processing, Vol. 19, No. 4, pp 371-551. 2025
机构 * Huawei Noah’s Ark Lab(华为诺亚实验室)
专题命中 跨模态检索 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.CL
Comments Paper accepted to NeurIPS 2025 DB
机构 * PitchBook USA(PitchBook美国公司)
专题命中 跨模态检索 :multimodal(title,abstract);image-text(abstract)
机构 * Huawei Technologies Co., Ltd.(华为技术有限公司)
专题命中 跨模态检索 :multimodal(title,abstract);分类 cs.CV、cs.CL、cs.AI
Comments Paper accepted to EMNLP-2025(Main)
专题命中 跨模态检索 :multi-modal(title);multimodal(abstract);分类 cs.CV、cs.AI
Comments Under review for ICRA 2026
机构 * Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学) ; Department of Mathematics and Computer Science, Faculty of Science, Alexandria University(亚历山大大学数学与计算机科学系) ; Sheikh Shakhbout Medical City(谢赫·沙赫布OUT医疗城)
专题命中 跨模态检索 :multimodal(title,abstract);分类 cs.CV、cs.CL
机构 * University of New South Wales and CSIRO’s Data61(新南威尔士大学和CSIRO的Data61)
专题命中 跨模态检索 :multimodal(abstract);multimodal foundation model(abstract);分类 cs.CL
机构 * Computer Science(计算机科学) ; Carnegie Mellon University(卡内基梅隆大学)
专题命中 跨模态检索 :multi-modal(abstract);分类 cs.AI
机构 * Shanghai Jiao Tong University(上海交通大学)
专题命中 多模态生成 :multi-modal(title,abstract);MLLM(abstract);分类 cs.CV、cs.CL
Comments Accepted to ACM MM 2025; 14 pages including Appendix
机构 * NVIDIA
专题命中 多模态生成 :multi-modal(title,abstract);分类 cs.AI
Comments Code and documentation are available here: https://github.com/isaac-sim/IsaacLab
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CL
Comments Published in IEEE Access
Journal ref IEEE Access, vol. 13, pp. 176751-176769, 2025
专题命中 多模态生成 :multimodal(abstract);MLLM(abstract);分类 cs.CL、cs.MM
Comments 17 pages, in Spanish language
机构 * Show Lab, National University of Singapore(新加坡国立大学Show实验室)
专题命中 多模态评测 :multimodal(title,abstract);分类 cs.CV、cs.CL、cs.AI
Comments Accepted by EMNLP 2025 Findings
专题命中 多模态评测 :multimodal(title,abstract);分类 cs.CV、cs.AI
机构 * Superb AI Seoul, South Korea(超霸AI首尔韩国)
专题命中 多模态评测 :multi-modal(title,abstract);分类 cs.CV、cs.AI
Comments 9 pages, 2 figures
机构 * DCST, Tsinghua University(清华大学信息国家实验室) ; Quan Cheng Laboratory(钱程实验室) ; AIR, Tsinghua University(清华大学人工智能研究院)
专题命中 多模态评测 :multimodal(abstract);分类 cs.CV、cs.MM
机构 * Krutrim AI ; OLA Electric
专题命中 多模态评测 :multimodal(abstract);分类 cs.CV、cs.AI
机构 * AI for Sensor Data Analytics Research Group(人工智能传感器数据解析研究组) ; Ulm University of Applied Sciences(乌尔姆应用科学大学) ; Biomechatronic Research Group(生物机械研究组) ; Institute of Computer Science(计算机科学研究所)
专题命中 多模态评测 :multimodal(abstract);分类 cs.CV、cs.AI
机构 * Department of Computer Engineering Isik University Istanbul, Turkey(计算机工程系伊斯肯大学) ; Department of Artificial Intelligence and Data Engineering(人工智能与数据工程系) ; Data Engineering Istanbul Technical University Istanbul, Turkey(数据工程系伊斯坦布尔技术大学)
专题命中 多模态评测 :multimodal(abstract);分类 cs.CV