Self-Enhanced Image Clustering with Cross-Modal Semantic Consistency
专题命中 图文多模态 :cross-modal(title,abstract);image-text(abstract);分类 cs.CV
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
专题命中 图文多模态 :cross-modal(title,abstract);image-text(abstract);分类 cs.CV
专题命中 图文多模态 :multimodal(title,abstract);分类 cs.CV、cs.CL、cs.MM
Comments 12 pages, 10 figures
Journal ref Proceedings of the Nineteenth International AAAI Conference on Web and Social Media, Vol. 19 (2025)
机构 * King Abdullah University of Science and Technology(卡斯特大学)
专题命中 图文多模态 :multimodal(title,abstract);分类 cs.CV、cs.CL
机构 * Dalian University of Technology(大连理工大学) ; Alibaba Group(阿里巴巴集团)
专题命中 图文多模态 :multi-modal(title,abstract);分类 cs.CV
Comments ACM Multimedia 2025
专题命中 图文多模态 :multimodal(abstract);cross-modal(abstract);分类 cs.CV、cs.CL
机构 * The University of Queensland(昆士兰大学) ; University of California, Merced(加州大学默塞德分校)
专题命中 图文多模态 :multimodal(abstract);cross-modal(abstract);分类 cs.CV、cs.AI
Comments Work in progress
专题命中 图文多模态 :cross-modal(abstract);image-text(abstract);分类 cs.CV
机构 * University of Science and Technology of China(中国科学技术大学) ; The Hong Kong Polytechnic University(香港理工大学) ; University of Washington(华盛顿大学) ; Nanjing University(南京大学) ; Stanford University(斯坦福大学) ; University of Chinese Academy of Sciences(中国科学院大学)
专题命中 图文多模态 :multimodal(abstract);image-text(abstract);分类 cs.CV
Comments ICCV 2025
专题命中 图文多模态 :multimodal(abstract);分类 cs.CV、cs.AI
机构 * Department of Intelligent Systems and Robotics, University of West Florida(智能系统与机器人系,西佛罗里达大学) ; Florida Institute For Human and Machine Cognition (IHMC)(佛罗里达人类与机器认知研究所) ; Center for Cybersecurity, University of West Florida(网络安全中心,西佛罗里达大学)
专题命中 图文多模态 :multimodal(abstract);分类 cs.CV、cs.AI
Journal ref Computers, Materials & Continua, 2025
机构 * Project Leader(项目负责人) ; Corresponding Author(通讯作者)
专题命中 图文多模态 :multimodal(abstract);分类 cs.CV、cs.AI
专题命中 图文多模态 :multimodal(abstract);分类 cs.CV、cs.AI
Comments 19 pages, 6 figures, 4 tables
机构 * Zeitview
专题命中 图文多模态 :multimodal(abstract);分类 cs.CV
Comments Accepted to 2025 ICCV VISION Workshop
机构 * Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
专题命中 图文多模态 :multimodal(abstract);分类 cs.CV
Comments This paper has been accepted by ACM Multimedia 2025 (ACM MM 2025)
机构 * School of Medicine, The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)医学院) ; Shenyang Institute of Automation, Chinese Academy of Sciences(中国科学院沈阳自动化研究所) ; Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院) ; Wave and Machine Intelligence Department, Technology Innovation Institute(技术创新研究院波浪与机器智能部门) ; University of the Chinese Academy of Sciences(中国科学院大学) ; McGill University(麦吉尔大学)
专题命中 图文多模态 :cross-modal(abstract);分类 cs.CV
Comments 10 pages
机构 * AI VIETNAM Lab(AI越南实验室) ; Carnegie Mellon University(卡内基梅隆大学) ; University of Wisconsin - Madison(威斯康星大学麦迪逊分校) ; University of Pittsburgh(匹兹堡大学) ; University of Alabama at Birmingham(阿拉巴马大学伯明翰分校) ; Northwestern University(西北大学)
专题命中 图文多模态 :image-text(abstract);分类 cs.CV
Comments 11 pages, 5 figures. Accepted to VisionDocs @ ICCV 2025
机构 * The Chinese University of Hong Kong, Department of Computer Science & Engineering(香港中文大学计算机科学与工程系)
专题命中 图文多模态 :image-text(abstract);分类 cs.CV
机构 * Intelligent Space Robotics Laboratory(智能空间机器人实验室) ; Skolkovo Institute of Science and Technology(斯克尔科沃科学与技术研究所)
专题命中 图文多模态 :multimodal(abstract)
机构 * Division of Information Science, Graduate School of Science and Technology, Nara Institute of Science and Technology(信息科学系,科学技术研究生学校,科学技术研究所) ; Department of Electrical and Electronic Engineering, Faculty of Engineering Science, Kansai University(电气电子工程系,工学科学大学) ; Department of Electronics, Kobe City College of Technology(电子系,神户市立技术学院) ; OMRON SINIC X Corporation(OMRON SINIC X公司)
专题命中 图文多模态 :multimodal(abstract)
Comments Published in IEEE Access, Jul 14 2025
Journal ref IEEE Access, vol. 13, pp. 125420-125441, 2025