Unlabeled Data Improves Fine-Grained Image Zero-shot Classification with Multimodal LLMs
未标记数据提升多模态大语言模型在细粒度图像零样本分类中的性能
Yunqi Hong, Sohyun An, Andrew Bai, Neil Y. C. Lin, Cho-Jui Hsieh
机构
*
Computer Science Department, University of California, Los Angeles(加州大学洛杉矶分校计算机科学系)
;
Mechanical and Aerospace Engineering Department, University of California, Los Angeles(加州大学洛杉矶分校机械与航空航天工程系)
机构
*
National University of Defense Technology(国防科技大学)
;
Shanghai AI Laboratory(上海人工智能实验室)
;
The Chinese University of Hong Kong(香港中文大学)
;
ShanghaiTech University(上海科技大学)
Leveraging Biomolecule and Natural Language through Multi-Modal Learning: A Survey
利用生物分子和自然语言通过多模态学习:一篇综述
Qizhi Pei, Zhimeng Zhou, Kaiyuan Gao, Jinhua Zhu, Yue Wang, Zun Wang, Tao Qin, Lijun Wu, Rui Yan
机构
*
Gaoling School of Artificial Intelligence, Renmin University of China(中国人民大学人工智能学院)
;
Zhejiang University(浙江大学)
;
Shanghai Innovation Institute(上海创新研究院)
;
Huazhong University of Science and Technology(华中科技大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Zhongguancun Academy(中关村学院)
;
Shanghai AI Laboratory(上海人工智能实验室)
;
School of Artificial Intelligence, Wuhan University(武汉大学人工智能学院)
HMR3D: Hierarchical Multimodal Representation for 3D Scene Understanding with Large Vision-Language Model
HMR3D:用于大视觉-语言模型的层次多模态表示以实现3D场景理解
Chen Li, Eric Peh, Basura Fernando
机构
*
Institute of High-Performance Computing, Agency for Science, Technology and Research, Singapore(高性能计算研究所,科技研究局,新加坡)
;
Centre for Frontier AI Research, Agency for Science, Technology and Research, Singapore(前沿人工智能研究中心,科技研究局,新加坡)
;
College of Computing and Data Science, Nanyang Technological University, Singapore(计算与数据科学学院,南洋理工大学,新加坡)
机构
*
Department of Clinical Neurosciences, University of Cambridge, UK(剑桥大学临床神经科学系)
;
Department of Electrical & Electronic Engineering, The University of Manchester, UK(曼彻斯特大学电气与电子工程系)
;
Department of Pathology, State Key Laboratory of Oncology in South China, Guangdong Provincial Clinical Research Center for Cancer, Sun Yat-sen University Cancer Center, China(南方医科大学肿瘤学国家重点实验室、广东省癌症临床研究中心、中山大学肿瘤中心病理学部)
;
Department of Statistics and Actuarial Science, The University of Hong Kong, Hong Kong SAR, China(香港大学统计与精算科学系)
;
Department of Clinical Neurosciences and Department of Applied Mathematics and Theoretical Physics, University of Cambridge(剑桥大学临床神经科学系和应用数学与理论物理系;邓迪大学科学与工程学院和医学学院)
;
School of Science and Engineering and School of Medicine, University of Dundee, UK
TinyRS-R1: Compact Multimodal Language Model for Remote Sensing
TinyRS-R1:用于遥感的紧凑多模态语言模型
Aybora Koksal, A. Aydin Alatan
机构
*
Center for the Image Analysis (OGAM) and Department of Electrical and Electronics Engineering of Middle East Technical University (METU)(图像分析中心(OGAM)和中欧技术大学(METU)电子与电气工程系)
CommentsAccepted to IEEE Geoscience and Remote Sensing Letters (GRSL). Code, models, and the captions for datasets are available at https://github.com/aybora/TinyRS
MQFL-FHE: Multimodal Quantum Federated Learning Framework with Fully Homomorphic Encryption
MQFL-FHE: 多模态量子联邦学习框架与全同态加密
Siddhant Dutta, Nouhaila Innan, Sadok Ben Yahia, Muhammad Shafique, David Esteban Bernal Neira
机构
*
SVKM's Dwarkadas J. Sanghvi College of Engineering(SVKM的达沃拉德·J·桑格维工程学院)
;
eBRAIN Lab, Division of Engineering, New York University Abu Dhabi (NYUAD)(eBRAIN实验室,工程系,纽约大学阿布扎比分校(NYUAD))
;
Center for Quantum and Topological Systems (CQTS), NYUAD Research Institute(量子与拓扑系统中心(CQTS),NYUAD研究机构)
;
The Maersk Mc-Kinney Moller Institute, University of Southern Denmark(马士克·麦克金尼·莫勒研究所,南丹麦大学)
;
Corvinus Institute for Advanced Studies (CIAS), Budapest, Hungary(科维努斯高级研究学院(CIAS),布达佩斯,匈牙利)
;
Davidson School of Chemical Engineering, Purdue University(戴维森化学工程学院,普渡大学)