Self-Improvement in Multimodal Large Language Models: A Survey
Shijian Deng, Kai Wang, Tianyu Yang, Harsh Singh, Yapeng Tian
机构
*
The University of Texas at Dallas(德克萨斯大学达拉斯分校)
;
University of Toronto(多伦多大学)
;
University of Notre Dame(诺特丹大学)
;
Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);分类 cs.CL
机构
*
State Key Laboratory of Novel Software Technology, Nanjing University(南京大学新型软件技术国家重点实验室)
;
School of Artificial Intelligence, Nanjing University(南京大学人工智能学院)
;
School of Intelligence Science and Technology, Nanjing University(南京大学智能科学与技术学院)
;
Department of Computer Science, ETH Zurich, Switzerland(瑞士苏黎世联邦理工学院计算机科学系)
;
School of Computing and Mathematical Sciences, Birkbeck, University of London, UK(伦敦大学伯克贝克学院计算与数学科学学院)
专题命中
评测与基准
:LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
Training-Free Out-Of-Distribution Segmentation With Foundation Models
Laith Nayal, Hadi Salloum, Ahmad Taha, Yaroslav Kholodov, Alexander Gasnikov
机构
*
Laboratory of Multimodal Research In Industry, AI Institute, Innopolis University(工业多模态研究实验室,人工智能研究所,因诺普利斯大学)
;
Phystech School of Applied Mathematics and Computer Science, Moscow Institute of Physics and Technology(物理与技术莫斯科应用数学与计算机科学学院,莫斯科物理技术学院)
;
Research Center for Artificial Intelligence, Innopolis University(人工智能研究中心,因诺普利斯大学)
;
Q Deep, Innopolis(Q深度,因诺普利斯)
;
Machine Learning and Data Representation Lab, Innopolis University(机器学习与数据表示实验室,因诺普利斯大学)
机构
*
State Key Laboratory of AI Safety, Institute of Computing Technology, Chinese Academy of Sciences(人工智能安全国家重点实验室,计算技术研究所,中国科学院)
;
Tongyi Lab, Alibaba Group(通义实验室,阿里巴巴集团)
;
Institute of Computing Technology(计算技术研究所)
机构
*
Amity Centre for Artificial Intelligence, Amity University, India(阿米蒂大学人工智能中心,阿米蒂大学,印度)
;
Amity School of Engineering and Technology, Amity University, India(阿米蒂工程与技术学院,阿米蒂大学,印度)
;
Amity Institute of Biotechnology, Amity University, India(阿米蒂生物技术学院,阿米蒂大学,印度)
A Survey of Pun Generation: Datasets, Evaluations and Methodologies
Yuchen Su, Yonghua Zhu, Ruofan Wang, Zijian Huang, Diana Benavides-Prado, Michael Witbrock
机构
*
School of Computer Science, University of Auckland(奥克兰大学计算机科学学院)
;
Singapore University of Technology and Design(新加坡科技设计大学)
;
School of Electronic Engineering and Computer Science, Queen Mary University of London(伦敦女王学院电子工程与计算机科学学院)
SpeechCT-CLIP: Distilling Text-Image Knowledge to Speech for Voice-Native Multimodal CT Analysis
Lukas Buess, Jan Geier, David Bani-Harouni, Chantal Pellegrini, Matthias Keicher, Paula Andrea Perez-Toro, Nassir Navab, Andreas Maier, Tomas Arias-Vergara
机构
*
Computer Aided Medical Procedures, Technical University of Munich(慕尼黑技术大学计算机辅助医学程序)