Libra-MIL: Multimodal Prototypes Stereoscopic Infused with Task-specific Language Priors for Few-shot Whole Slide Image Classification
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.AI
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.AI
专题命中 多模态训练与对齐 :multimodal(title,abstract);image-text(abstract);分类 cs.CL、cs.MM
机构 * GIFT University(GIFT大学) ; La Trobe University(拉特罗布大学) ; Department of Primary Industries(初级产业部门)
专题命中 多模态训练与对齐 :cross-modal(title,abstract);multi-modal(abstract);分类 cs.CV、cs.AI
Comments Accepted at 26th International Conference on Digital Image Computing: Techniques and Applications (DICTA 2025)
机构 * EPFL(瑞士联邦理工学院) ; University of Basel(巴塞尔大学) ; HSLU(苏黎世联邦理工学院)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.AI
Comments NeurIPS 2025 camera-ready
机构 * Sanya Science and Education Innovation Park, Wuhan University of Technology, Sanya 572025, China(武汉理工大学三亚科学教育创新园) ; School of Computer Science and Artificial Intelligence, Wuhan University of Technology, Wuhan 430070, China(武汉理工大学计算机科学与人工智能学院) ; Key Laboratory of Ocean Circulation and Waves, Institute of Oceanology, Chinese Academy of Sciences, Qingdao 266071, China(中国科学院海洋循环与波浪重点实验室) ; Hubei Key Laboratory of Transportation Internet of Things, School of Computer Science and Artificial Intelligence, Wuhan University of Technology, Wuhan 430070, China(湖北省交通物联网重点实验室) ; State Key Laboratory of Maritime Technology and Safety, Wuhan University of Technology, Wuhan 430063, China(武汉理工大学航海技术与安全国家重点实验室) ; College of Oceanography and Ecological Science, Shanghai Ocean University, Shanghai 201306, China(上海海洋大学海洋科学与生态学院) ; School of Computer and Artificial Intelligence, Zhengzhou University, Zhengzhou 450001, China(郑州大学计算机与人工智能学院) ; Rapid-Rich Object Search Lab, School of Electrical and Electronic Engineering, Nanyang Technological University, Singapore 639798(南洋理工大学电子与电气工程学院)
专题命中 多模态训练与对齐 :multi-modal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.AI
Comments 14 pages, 8 figures
机构 * Zhejiang Gongshang University(浙江工商大学) ; Ningbo Key Laboratory of Spatial Intelligence and Digital Derivative(宁波空间智能与数字衍生关键实验室) ; Institute of Digital Twin, Eastern Institute of Technology, Ningbo(数字孪生研究院,东部技术研究所,宁波) ; Meituan Inc.(美团公司) ; National University of Singapore(新加坡国立大学)
专题命中 多模态训练与对齐 :multimodal(title,abstract);MLLM(abstract);分类 cs.CV、cs.AI
Comments 9 pages, 6 figures, accepted by EMNLP2025
机构 * Artificial Creative intelligence (ACI)(人工创意智能(ACI)) ; Expedia Group(Expedia集团)
专题命中 多模态训练与对齐 :multi-modal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.AI
机构 * University of Bristol(布里斯托大学) ; Brown University(布朗大学) ; South China University of Technology(华南理工大学)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.CL
Comments Accepted by conference EMNLP2025
机构 * Jadavpur University(贾瓦帕尔大学)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.AI
Comments Accepted at the IEEE/CVF International Conference on Computer Vision (ICCV 2025), Workshop on Curated Data for Efficient Learning
机构 * Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所) ; Institute of AI for Industries(工业人工智能研究所) ; University of Chinese Academy of Sciences(中国科学院大学)
专题命中 多模态训练与对齐 :multimodal(title,abstract);MLLM(abstract);分类 cs.CV、cs.AI
机构 * Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所) ; Institute of AI for Industries(工业人工智能研究所)
专题命中 多模态训练与对齐 :multimodal(title,abstract);MLLM(abstract);分类 cs.CV、cs.AI
Comments 5 pages, 5 figures
机构 * Digital Medical Research Center, School of Basic Medical Science, Fudan University, Shanghai 200032, China(复旦大学基础医学学院数字医学研究中心) ; Shanghai Key Lab of Medical Image Computing(上海医学图像计算重点实验室)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.AI
机构 * Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) ; Imperial College London(帝国理工学院伦敦分校) ; The Hong Kong University of Science and Technology(香港科学与技术大学)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.AI
Comments Early accepted by MICCAI2025
机构 * Singapore University of Social Sciences(新加坡社会科学研究大学)
专题命中 多模态训练与对齐 :multimodal(title,abstract);multi-modal(abstract);分类 cs.CV、cs.AI
机构 * Zhejiang University(浙江大学)
专题命中 多模态训练与对齐 :MLLM(title,abstract);multimodal(abstract);分类 cs.CV、cs.AI
机构 * Khoury College of Computer Sciences, Northeastern University(东北大学克劳尔计算机科学学院) ; New Laboratory of Pattern Recognition, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所模式识别新实验室)
专题命中 多模态训练与对齐 :cross-modal(title,abstract);multi-modal(abstract);分类 cs.CV、cs.AI
机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) ; The Hong Kong University of Science and Technology(香港科技大学) ; Technical University of Munich(慕尼黑技术大学) ; Columbia University(哥伦比亚大学)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CL、cs.AI
Comments 9 pages, 6 figure, 3 tables
Journal ref IJCAI 2025 main conference
机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) ; School of Computer Science and Engineering, Sun Yat-Sen University(中山大学计算机科学与工程学院)
专题命中 多模态训练与对齐 :multimodal(title,abstract);multi-modal(abstract);分类 cs.CV、cs.AI
Comments ACM MM 2025
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.AI
Comments 9 pages,7 figures,conference
机构 * Brown University(布朗大学) ; University of Warwick(沃里克大学) ; Samsung US(三星美国分公司) ; Boston University(波士顿大学) ; University of Alabama at Birmingham(阿拉巴马大学伯明翰分校) ; Rutgers University(罗格斯大学)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.AI
Comments COLM 2025, 30 pages, 10 figures, 16 tables
机构 * Fujitsu Research of Europe(富士通欧洲研究)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.AI
Comments International Conference on Machine Learning (ICML) 2025 (Spotlight)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CL、cs.AI
Comments arXiv admin note: This paper has been withdrawn by arXiv due to disputed and unverifiable authorship
机构 * University of Wisconsin-Madison(威斯康星大学麦迪逊分校) ; Purdue University(普渡大学) ; The University of Hong Kong(香港大学)
专题命中 多模态训练与对齐 :multimodal(title,abstract);MLLM(abstract);分类 cs.CV、cs.AI
Comments Published at ICCV 2025
机构 * School of Computing and Artificial Intelligence(计算机与人工智能学院) ; Southwest Jiaotong University(西南交通大学) ; Department of Gastroenterology(消化内科部) ; The Third People’s Hospital of Chengdu(成都第三人民医院)
专题命中 多模态训练与对齐 :cross-modal(title,abstract);image-text(abstract);分类 cs.CV、cs.AI
机构 * Tianjin University of Technology(天津理工大学) ; Hainan University(海南大学)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.AI
Comments This paper has been accepted by ACM MM 2025
机构 * Department of Computer Science and Engineering, Bangladesh University of Business and Technology(计算机科学与工程系,孟加拉国商业技术大学) ; Department of Computer Science, American International University-Bangladesh(计算机科学系,美国国际大学-孟加拉国)
专题命中 多模态训练与对齐 :multi-modal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.CL
机构 * Department of Computer and Information Science, University of Macau(计算机与信息科学系,澳门大学)
专题命中 多模态训练与对齐 :multi-modal(title,abstract);multimodal(abstract);分类 cs.CV、cs.AI
专题命中 多模态训练与对齐 :multi-modal(title,abstract);multimodal(abstract);分类 cs.CV、cs.CL
Comments Accepted by IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI), 2025
专题命中 多模态训练与对齐 :multimodal(title,abstract);image-text(abstract);分类 cs.CV、cs.AI
机构 * School of Artificial Intelligence, Anhui University(安徽大学人工智能学院) ; Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(合肥国家综合科学中心人工智能研究所) ; Anhui Province Key Laboratory of Affective Computing and Advanced Intelligent Machines, School of Computer Science and Information Engineering, Hefei University of Technology(安徽省情感计算与先进智能机器重点实验室,合肥工业大学计算机科学与信息工程学院) ; Department of Computer and Network Engineering, The University of Electro-Communications(电子通信大学计算机与网络工程系)
专题命中 多模态训练与对齐 :cross-modal(title,abstract);image-text(abstract);分类 cs.CV、cs.AI
Comments Accepted by IEEE TIP